Why
The PoC runs the same task under zero-shot, few-shot and formatted prompts against a fixed eval set, turning prompt choices into a measured comparison.
How it works
Not yet built.
Workspace Index › Dev Notes › Prompt engineering — the few-shot examples do most of the work
#192PoC
How you phrase and exemplify a task often changes accuracy more than which model you pick, and the discipline is measuring that rather than trusting intuition about wording.
The PoC runs the same task under zero-shot, few-shot and formatted prompts against a fixed eval set, turning prompt choices into a measured comparison.
Not yet built.
과제를 어떻게 표현하고 예시하느냐가 어느 모델을 고르느냐보다 정확도를 더 바꾸는 경우가 많고, 이 분야의 규율은 문구에 대한 직관을 믿는 대신 그것을 측정하는 것입니다.
이 PoC는 고정된 평가 집합에 대해 같은 과제를 제로샷·소수샷·서식 프롬프트로 돌려, 프롬프트 선택을 측정된 비교로 바꿉니다.
아직 만들지 않음.