Policy · February 17, 2026

IB internal assessment was built on a conversation. That helps right now

Viva Team · 5 min read

A folder whose contents are a conversation held over time
A folder whose contents are a conversation held over time
정책 · 2026년 2월 17일

IB 내부 평가는 원래 대화 위에 세워졌어요. 지금 그게 도움이 돼요

Viva 팀 · 5분 분량

시간을 두고 이어진 대화가 담긴 폴더
시간을 두고 이어진 대화가 담긴 폴더

Something worth noticing in the current panic: the International Baccalaureate has required teachers to confirm the authenticity of internal assessment work for a very long time, and the mechanism it specifies is not software. It is the supervisor knowing the student's work well enough to vouch for it, through scheduled check-ins during the process.

The Extended Essay makes this explicit with the reflection sessions and the viva voce at the end. A supervisor is expected to talk to the student about their own research and form a judgment.

That design is not a response to generative AI. It predates it by decades. It happens to be the right shape.

Why the shape holds up

The IB model assumes the artefact alone is insufficient evidence, and that the relationship between a supervisor and the work over time is what authenticates it. Both assumptions survived generative models untouched, which cannot be said for most assessment design of the same era.

The parts of assessment that assumed the document proves the learning are the parts in trouble. The parts that assumed a person has to vouch for it are fine.

Where it strains

It strains on load. The model works because a supervisor carries a small number of students through a long process. Scale that to every subject, every internal assessment, in a school where teachers already supervise a full timetable, and the check-ins compress into a signature.

It also strains on consistency. Two supervisors in the same school can hold noticeably different bars, and the framework relies on professional judgment rather than a common script. That is a strength when supervisors have time and a weakness when they do not.

What schools are doing about it

The IB schools we work with are not replacing supervisor judgment. They are trying to make the routine part cheaper so the judgment part still happens.

None of that replaces the viva. It makes the viva better informed, which is the direction any change here should run.

The uncomfortable observation for everyone else: the framework that is coping best right now is the one that never fully believed the document in the first place.

지금의 소동 속에서 눈여겨볼 만한 게 있어요. 국제 바칼로레아(IB)는 아주 오래전부터 교사가 내부 평가 작업의 진정성을 확인하도록 요구해 왔고, 그 방법으로 지정한 건 소프트웨어가 아니에요. 진행 중에 예정된 면담을 통해, 지도교사가 학생의 작업을 보증할 만큼 잘 아는 것이에요.

확장 에세이(EE)는 이걸 성찰 세션과 마지막 구술 확인으로 명시해 놓았어요. 지도교사는 학생과 그 학생의 연구에 대해 이야기하고 판단을 내리도록 되어 있어요.

이 설계는 생성형 AI에 대한 대응이 아니에요. 수십 년 앞서 있었어요. 그런데 마침 모양이 맞아요.

이 모양이 버티는 이유

IB 모델은 결과물만으로는 근거가 부족하다고 가정하고, 지도교사와 작업 사이에 시간을 두고 쌓인 관계가 진정성을 확인해 준다고 봐요. 두 가정 모두 생성형 모델을 그대로 통과했어요. 같은 시기의 다른 평가 설계에 대해서는 하기 어려운 말이에요.

문서가 학습을 증명한다고 가정한 부분이 곤경에 빠졌어요. 사람이 보증해야 한다고 가정한 부분은 멀쩡해요.

무리가 오는 지점

부하에서 무리가 와요. 이 모델은 지도교사가 소수의 학생을 긴 과정 동안 데리고 갈 때 작동해요. 모든 과목, 모든 내부 평가로 확장하고, 교사가 이미 꽉 찬 시간표를 감당하는 학교에 놓으면, 면담은 서명으로 압축돼요.

일관성에서도 무리가 와요. 같은 학교의 두 지도교사가 눈에 띄게 다른 기준을 가질 수 있고, 프레임워크는 공통 대본이 아니라 전문가의 판단에 기대요. 시간이 있을 때는 강점이고 없을 때는 약점이에요.

학교들이 하고 있는 것

저희와 일하는 IB 학교들은 지도교사의 판단을 대체하지 않아요. 반복적인 부분을 싸게 만들어서 판단하는 부분이 실제로 일어나게 하려는 거예요.

이 중 무엇도 구술 확인을 대체하지 않아요. 구술 확인이 더 많은 정보를 갖게 만들 뿐이고, 여기서의 변화는 그 방향으로 가야 해요.

다른 이들에게 불편한 관찰 하나. 지금 가장 잘 버티고 있는 프레임워크는, 애초에 문서를 온전히 믿지 않았던 프레임워크예요.

See what an oral check actually looks like.

A real interview and the report an instructor receives. Three minutes, no sign-up.

구술 확인이 실제로 어떤 모습인지 보세요.

실제 인터뷰와 선생님이 받는 리포트예요. 3분이면 되고 가입도 필요 없어요.