Practice · September 17, 2025

Group projects have always had a free rider problem. AI made it invisible

Viva Team · 5 min read

Four names on one document, three lines attached
Four names on one document, three lines attached
실무 · 2025년 9월 17일

조별 과제에는 늘 무임승차 문제가 있었어요. AI가 그걸 안 보이게 만들었어요

Viva 팀 · 5분 분량

문서 하나에 이름 넷, 이어진 선은 셋
문서 하나에 이름 넷, 이어진 선은 셋

Every instructor who has assigned group work knows the shape of the complaint. Four names on a document, two people wrote it, and the peer evaluation forms come back saying everyone contributed equally because nobody wants to be the person who sank a classmate.

That was true before generative models. What changed is the cover story. "I did the literature review section" used to require doing a literature review. It is now a sentence anyone can say about a section they did not write, produced by a group member who also did not write it.

Why peer evaluation underperforms

The instrument has a well-known set of problems: friendship inflation, retaliation fear, and a strong tendency for everyone to award the middle score. Weighted variants and confidential forms help at the margin. None of them fix the underlying issue, which is that you are asking students to make a social judgment and then treating the result as a measurement.

You are not trying to find out who students think worked hardest. You are trying to find out who can account for the work.

The more direct question

Ask each member to explain the part they say they did. Not to justify their teammates, not to allocate credit, just to talk through their own contribution: what they chose, what they tried that did not work, why the section is structured the way it is.

This is unusually informative, and it has a property peer evaluation lacks: the student is being asked about themselves, so a bad outcome is not a betrayal of anyone. Free riders reveal themselves without a classmate having to name them, which is precisely the dynamic that breaks peer forms.

It also surfaces something more useful than a free rider list. When two students both claim the same component and one of them can describe the decisions inside it and the other cannot, you have learned something specific. When a quiet student turns out to be the one who can explain the analysis, you have learned something better.

Keeping it proportionate

A few practical notes from instructors doing this.

The unglamorous truth is that group work was always assessed on trust, and trust worked reasonably well when producing the artefact required doing the work. That assumption is what broke. The response does not need to be more forms.

If you want the mechanics of how we do this, it is written up at group projects.

조별 과제를 내 본 선생님이라면 이 불만의 모양을 알아요. 문서에 이름 넷, 실제로 쓴 사람은 둘, 그리고 동료 평가지는 모두가 똑같이 기여했다고 돌아와요. 아무도 동기를 가라앉힌 사람이 되고 싶지 않으니까요.

생성형 모델 이전에도 그랬어요. 달라진 건 변명의 설득력이에요. "저는 선행연구 부분을 맡았어요"는 예전에는 실제로 선행연구를 해야 할 수 있는 말이었어요. 지금은 쓰지 않은 부분에 대해 누구나 할 수 있는 문장이고, 그 부분을 만든 조원도 쓰지 않았을 수 있어요.

동료 평가가 힘을 못 쓰는 이유

이 도구에는 잘 알려진 문제들이 있어요. 친분에 따른 점수 인플레이션, 보복에 대한 두려움, 그리고 모두에게 중간 점수를 주려는 강한 경향이요. 가중치를 두거나 익명으로 하면 가장자리에서는 나아져요. 하지만 근본 문제는 그대로예요. 학생에게 사회적 판단을 요구해 놓고 그 결과를 측정치처럼 다루는 거니까요.

학생들이 누가 가장 열심히 했다고 생각하는지를 알아내려는 게 아니에요. 누가 그 작업을 설명할 수 있는지를 알아내려는 거예요.

더 직접적인 질문

각자 자기가 맡았다고 한 부분을 설명하게 하세요. 조원을 변호하라는 것도, 공을 배분하라는 것도 아니고, 그냥 자기 기여를 이야기하는 거예요. 무엇을 선택했고, 해 봤는데 안 된 게 무엇이고, 왜 그 부분이 그런 구조인지요.

이건 정보량이 유난히 많고, 동료 평가에 없는 성질이 하나 있어요. 학생이 자기 자신에 대해 질문받으니까, 결과가 나빠도 누군가를 배신한 게 아니에요. 무임승차는 동기가 이름을 대지 않아도 드러나요. 동료 평가지를 망가뜨리는 게 바로 그 이름 대기거든요.

무임승차자 명단보다 쓸모 있는 것도 드러나요. 두 학생이 같은 부분을 자기 것이라 하는데 한 명은 그 안의 결정들을 설명하고 다른 한 명은 못 한다면, 구체적인 걸 알게 된 거예요. 조용한 학생이 분석을 설명할 수 있는 사람으로 밝혀진다면 더 좋은 걸 알게 된 거고요.

과하지 않게 유지하기

이걸 하고 있는 선생님들의 실무 메모예요.

멋없는 진실은, 조별 과제는 늘 신뢰 위에서 평가돼 왔고 결과물을 만들려면 실제로 일을 해야 했던 동안에는 그 신뢰가 꽤 잘 작동했다는 거예요. 무너진 건 그 가정이에요. 대응이 서류를 더 늘리는 것일 필요는 없어요.

저희가 이걸 어떻게 구현했는지는 조별 과제에 적어 두었어요.

See what an oral check actually looks like.

A real interview and the report an instructor receives. Three minutes, no sign-up.

구술 확인이 실제로 어떤 모습인지 보세요.

실제 인터뷰와 선생님이 받는 리포트예요. 3분이면 되고 가입도 필요 없어요.