What the Gathering Effectiveness Score actually measures.
In the AI age, broadcast information is cheap. What matters is whether a gathering creates what AI cannot fake for a group: commitment, trust, transfer, useful ties, and verifiable movement. Scroll the story; the panel on the right shows the evidence behind each step.
- A satisfaction score or a show of hands
- Applause and a good feeling in the room
- Tells you how it felt, not what changed
- One 0 to 100 score across the 8 things that make a gathering work
- Decisions made, skills used, relationships built, follow-through that sticks
- Shows what worked and the one thing to fix
Satisfaction is no longer enough.
A room can feel good and still fail to create commitment, transfer, trusted ties, or measurable movement. The grade has to move from feeling to proof.
The score starts with eight mechanisms.
Gatherings are systems. Participation, follow-through, focus, fit, networks, transfer, evidence, and future readiness each explain a path from event to outcome.
The smile sheet is a weak signal.
Enjoyment, room feel, and speaker polish correlate with measured learning at about r=.02, almost no relationship. Usefulness is the only self-report worth keeping.
Replace reaction with an outcome ladder.
Attendance, activity, and perception are useful context, but they sit below proof. The real test is competence, transfer, behavior, and results.
Passive broadcast fails the job.
Randomized evidence shows active formats can produce more learning even when people feel they learned less. That is why participation is a pillar, not a style preference.
Evidence maturity protects the claim.
Baseline, comparison, isolation, and follow-up turn a claim into something a hostile reviewer could verify. No baseline, no proven value.
Change the gathering. Change the measure.
The mission is not better surveys. It is to design gatherings that create movement, then measure owners, dates, behavior, transfer, relationships, and results honestly.
What the score actually grades.
Each pillar scores 0 to 100. Together they measure the mechanisms most likely to produce outcomes, plus the evidence maturity needed to prove them.
Participation architecture
People doing, not just watching: the share of the agenda that is interactive.
People doing vs watching.Follow through
Named owners, dates, commitments, and visible tracking after the room.
Intentions converted to action.Problem specificity
One costly, named problem the gathering is actually working on.
The work has a target.Personalization
Content and connections tailored to attendee goals, not one-size-fits-all.
Usefulness beats enjoyment.Network design
Deliberate bridging so the right people meet, instead of by chance.
The right people connect.Learning transfer
Applied behavior at 30 to 90 days, not a feeling at the exit door.
Learning survives the room.Evidence maturity
Baseline, comparison, and follow-up so a skeptic could verify what changed.
The claim can be checked.Future-of-work fit
Value created against time, energy, hybrid reality, and AI-augmented work.
Fits how work happens now.The rigorous version.
What counts as a gathering +
Any convening of 25 or more people with a stated purpose. Past about 25 people the room stops self-correcting, so the agenda becomes the operating system.
Why satisfaction is rejected +
Meeting science treats satisfaction as an attitude, not an outcome. Affective reaction correlates with measured learning at about r=.02. We record satisfaction only as context; it never drives the grade.
How a score is computed +
Each of the eight pillars is read from the public agenda and any linked evidence, scored 0 to 100, then rolled up. Where possible, outcomes are measured at a lag: 30 to 90 days for behavior, months for relationships and results.
Honesty about evidence strength +
More than half the evidence rows in the corpus are model inference, not read directly from source text, and every record says so. A missing signal in a public agenda is not proof the event lacked it.