Conrad Challenge Archive 2024 — 2026

The Judging Playbook — What Actually Happens to Your Submission

Derived from: the 2024-25 Innovation Stage Judge Guide, the Scoring Guide, the Chief Judge Q&A (Simon Glinsky), and close reading of five real Innovation Briefs assigned for judging in 2025 (Greencrete, SUNSYNC, Ocean Energy Dynamics, P-Bump, Puppy WC — all Energy & Environment, Global chapter).

This document exists to answer one question: what separates a 2 from a 4?


Part I — The judge's actual situation

Internalise this before writing anything.

Your judge is a volunteer professional — an entrepreneur, VC, engineer, scientist, or educator. They have 5 teams and 5–10 total hours. That is ~90 minutes for you, including writing five prose comment fields. They are explicitly instructed to:

  • Google your claims. "Perform an online search to verify originality of the approach or innovation."
  • Read the whole submission before scoring, and to re-read across multiple sessions.
  • Watch the video and open the website.
  • Open your references PDF.
  • Score each theme independently — a broken innovation can still earn 4 on Marketing.
  • Write as a coach, not a grader.
  • Never contact you. Every question they have goes unanswered — and becomes a doubt.

The five judge questions, in the order they get asked

  1. (Innovation) Have I seen this before? → they search
  2. (Practicality) Would this actually work? → they apply physics/domain sense
  3. (Storytelling) Do I believe these people? → they read for polish and coherence
  4. (Marketing) Do they know who buys this? → they look for named customers and named competitors
  5. (Finances) Do the numbers add up? → they check arithmetic, leniently

The single most important consequence: because the judge cannot ask you anything, every ambiguity resolves against you. "I am left wondering" is the phrase that costs points.


Part II — The five briefs, scored and diagnosed

I read these as an assigned judge. Names as submitted. Scores below are my judgments applying the official rubric, not official Conrad scores.

Calibration table

Brief Innov (30) Story (20) Pract (20) Mkt (20) Fin (10) ~Total Verdict
Ocean Energy Dynamics 3 5 2 5 4 ~76 Superb business writing, physics problem
P-Bump 3 4 2 4 3 ~66 Real engineering, fatal energy-accounting error
SUNSYNC 2 3 4 3 3 ~58 Real prototype, unoriginal innovation
Puppy WC 1 3 2 3 2 ~42 Fluent prose, product already exists
Greencrete 1 1 1 1 1 ~20 Unfinished

None of these is a finalist. Finalists need ~85–90+. Note what does not correlate with the total: word count, prose fluency, and enthusiasm.


Case 1 — Greencrete (~20/100): the anatomy of a 1

Low-carbon concrete using biomimicry. Two students.

What a judge sees immediately:

  • Q2 (Team) describes skiing and basketball hobbies and never states a role. 150 words spent on nothing.
  • Q7 (Competition) answers the wrong question — it repeats target-customer text instead of naming competitors.
  • Q8 (Go-to-Market) is copy-pasted verbatim from Q6.
  • Q9 says the business model is "listening" — a typo for licensing, uncorrected.
  • Q5 (Validation) opens: "we don't really have a complete 100% solution."
  • Q10 (Fundraising): total development cost "5–10 thousand dollars" to commercialise a new cement chemistry.
  • Innovation image: a Tinkercad tree on a brown box.
  • Website: a raw Canva /edit link — a judge clicking it may land in an editor, not a site.

Diagnosis: this is not a bad idea, it is an unfinished submission. Low-carbon concrete is a legitimate multi-billion-dollar problem; Holcim is correctly named. The team had a real thread (biomimicry → spiral fibre reinforcement) and abandoned it.

What the Chief Judge says to do here: score it honestly — 1s and 2s across the board — but comment as a coach. "It's reasonable for this team to receive one or two-point scores... That fulfills goal #2. Our #1 goal is to create learning for the students, including those who underperform."

Teaching value: highest of the five. Every failure is mechanical and fixable in an afternoon: answer the question asked, don't duplicate answers, proofread, publish the website properly.


Case 2 — Puppy WC (~42/100): fluent writing cannot rescue a non-innovation

A self-cleaning, flushable pet toilet with sensors. Two students, South Korea.

The prose is genuinely the second-best of the five: clean paragraphs, a systematic competitive comparison against puppy pads, washable pads, and artificial grass mats.

And it scores 1 on Innovation, because the judge is required to run an originality search — and automatic self-cleaning flushing pet toilets are an existing consumer product category (Inubox, BrilliantPad, and others). The brief never mentions that these exist. It positions only against pads and mats, which is the comparison set that makes the product look novel.

Compounding it:

  • No engineering specificity anywhere — no sensor type, no water volume per flush, no power draw, no cost.
  • "Proprietary flushing mechanism... engineered to accommodate the specific consistency of pet waste" — asserted, never described.
  • Quantification is a range with no basis: "hundreds to over a thousand disposable pads per year."

The lesson, and it is the single most important one in this document:

Choosing a favourable comparison set is the most common way teams destroy their own Innovation score. The judge picks the comparison set, not you. If a competitor exists and you don't name it, the judge concludes you either didn't look or you're hiding it. Both are worse than the competitor.

The fix is not a different product — it is Q7 naming the real incumbents and then earning the differentiation honestly ("existing units cost $500+ and require proprietary cartridges; ours...").


Case 3 — SUNSYNC (~58/100): a working prototype with nothing new in it

Solar panel that tracks the sun (photoresistors + servo + Arduino) and concentrates light with convex lenses. Three students, City of Knowledge Academy, Nigeria.

This team did more real work than any other in the set and scores mid-table. That is the lesson.

Strengths a judge genuinely rewards:

  • A built, tested prototype. Practicality = 4.
  • Real validation prose: they describe covering one light sensor to verify servo response, and checking the 9 V panel against the Arduino's 5 V/3.3 V logic to avoid damage. That is authentic engineering process, and it reads as authentic.
  • Clean structure, cited sources, clear roles, SDG 7 framing.

Why Innovation = 2:

  • Solar tracking is 1960s technology; concentrated photovoltaics is a mature field. Both are the first hits of the originality search the judge is required to run.
  • The combination is not defended as novel — no argument for why tracker + lens together is more than the sum.
  • "We used a unique method of programming the Arduino board so our code can't be replicated." This one sentence is actively damaging. To an engineer-judge it signals a fundamental misunderstanding of software, of reverse engineering, and of what IP is. IP Defensibility is an explicit sub-criterion; this answers it wrongly and confidently.
  • "Generates twice as much energy as normal solar panels, up to 90%" — two incompatible claims in one sentence, no measurement, despite the team owning a working rig that could have measured it.

Finances: $120 prototype → $1,200 commercial unit, with no bridge explaining the 10×. "$500 for the initial phase" of rollout is not a rollout budget.

Diagnosis: they had the instrument to answer the question and didn't run the experiment. A single afternoon logging watt-hours from a tracked+lensed panel versus a fixed panel, plotted, would have moved Innovation and Practicality both — and turned a marketing claim into evidence.

The SUNSYNC rule: if you built it, measure it. An unmeasured prototype is worth less than a well-reasoned design, because it proves you had access and chose not to look.


Case 4 — P-Bump (~66/100): rigorous arithmetic, wrong physics

Piezoelectric speed bump: hydraulic 10:1 force amplification onto a PZT stack, for Indonesia. Three students.

The most technically ambitious brief of the five. Genuine specifications: 10 cm → 31.6 cm cylinder bore (correctly giving 10:1 area ratio), a 4-section PZT stack of 100 layers each, 28×28 cm base, vacuum return mechanism, inline URLs to ScienceDirect and traffic-authority data, a named real competitor (the Shibuya Station piezoelectric floor) with honest advantages and disadvantages, a utility-patent + trade-secret strategy, and a coherent B2G model with Power Purchase Agreements.

This is what a 4–5 on Marketing and Practicality is supposed to look like structurally.

And Practicality still scores 2, for one reason a judge in this field will catch in 30 seconds:

A speed bump does not harvest free energy. It harvests energy from the vehicle's engine.

Depressing the bump does work on the vehicle — extra rolling resistance, extra fuel burned, at internal-combustion efficiency (~25%) and then piezo conversion. The device is a net energy loss machine that quietly taxes every driver. The brief never mentions this. Every credible piezo-road pilot has failed on exactly this economics.

Then the arithmetic, which looks rigorous and isn't:

  • Claimed stress on the PZT stack: 46,875 Pa. That is ~0.05 MPa. PZT is typically driven at tens of MPa. The stack is loaded at roughly a thousandth of a useful stress — the geometry defeats the amplification.
  • Energy per car computed as E = P·V with a 1 mm displacement → 4,113 J per car, then scaled by 20,000 cars/day to "power a household." The 1 mm displacement of a 28×28 cm ceramic stack is not physically available at that stress, and the formula conflates pressure-volume work with recoverable electrical energy at 88% "efficiency" of the material.

The lesson — and it is the opposite of the Puppy WC lesson:

Numbers do not create credibility. Numbers create checkable claims. A specific wrong number is worse than an honest range, because it is falsifiable and the judge is instructed to falsify it.

The redemption path was available and cheap: a sanity-check paragraph. "Energy is drawn from the vehicle, so this is only viable where the bump is already required for traffic calming and the marginal fuel cost is accepted as the price of the safety feature." That reframing is honest, survives scrutiny, and is more interesting.


Case 5 — Ocean Energy Dynamics (~76/100): the best business writing, the worst physics

Wave-energy system integrated into a ship's hull: one-way inlet, debris mesh, turbulence-reducing grids, lead oscillators on slide rods, sealed hydraulic cylinders, carbon-fibre hydraulic turbines. Four students.

This is what finalist-grade Marketing, Finance and Storytelling writing looks like. Study these:

  • Named real competitors: CalWave, Mocean Energy, Carnegie Clean Energy — and a precise positioning gap ("fixed-location systems... entirely unsuitable for moving vessels").
  • Segmented customers with a payer/buyer distinction: yacht owners (primary), shipbuilders (secondary), charter companies (tertiary), defence (adjacent) — each with its own reason to buy.
  • Sized market with a trajectory: $240 B (2023) → $460 B (2033) shipbuilding.
  • Real unit pricing with tiers: $50,000 single-unit for small yachts; 100, 000–250,000 multi-unit.
  • Recurring revenue: maintenance contracts, licensing, premium monitoring.
  • Use of funds as percentages: $10 M split 20% R&D / 40% manufacturing / 20% pilots / 20% marketing.
  • Layered IP strategy: patents on grids and hydraulic loop, trade secrets on manufacturing and material composition, trademarks for brand.
  • Honest disadvantage stated: higher initial cost, then answered with segment logic (wealthy, image-conscious buyers).
  • Attachments used properly: a 3D model, a rendered animation, and a partial physical prototype (turbine + DC motor + Arduino) with the limitation stated plainly ("we could not use water because the motor was not waterproof").

That last point deserves emphasis: they stated what their prototype could not do. Judges reward this. It is the opposite of SUNSYNC's unmeasured 90% claim.

So why does Practicality score 2 and Innovation only 3?

A ship's own motion in waves is the energy source. Extracting energy from that motion adds drag and increases the ship's propulsion demand. The brief then states the recovered electricity "powers an innovative propulsion system" — which closes the loop into a perpetual motion machine.

A marine-engineering judge sees this instantly. There is a legitimate version of this idea — parasitic harvesting from wave motion for hotel loads on a moored or drifting vessel, where the drag penalty is irrelevant because the ship isn't going anywhere. That version is defensible, still novel-ish, and would have scored 4 on Practicality with the same writing.

Diagnosis: the team out-wrote their engineering review. No adult with a fluid-dynamics background read Q4 before submission. Everything else was finalist-grade.


Part III — The cross-cutting patterns

The four failure archetypes

Archetype Symptom Example Cost
Unfinished Duplicated answers, typos, wrong question answered, broken links Greencrete Everything
Favourable comparison set Positions only against weak alternatives; real incumbent unnamed Puppy WC Innovation → 1
Unmeasured prototype Built the thing, never quantified it; claims unbacked SUNSYNC Innovation, Practicality
Polish over physics Excellent business writing wrapping an unsound mechanism Ocean Energy, P-Bump Practicality → 2, caps total ~75

The last one is the most dangerous, because it is invisible to the team, to a business-trained coach, and to any AI writing assistant. It is only caught by a domain expert reading for mechanism.

The Innovation score is decided by one paragraph

Innovation is 30% — the largest single block. The Chief Judge's stated test:

"If anyone can copy that application, it's not a strong innovation."

And the worked example: a new soda flavour scores 1/5. The same new flavour, if it reduces tooth decay, scores 3–5 — "a worthy and protectable innovation, even though it's a new application of an existing product."

So Innovation is not measured by technological sophistication. It is measured by defensible differentiation. SUNSYNC used more advanced hardware than the soda example and scored 2.

The sub-criteria say this outright: Verification (did an online search rule out duplicates?) and IP Defensibility (patent, trade secret, copyright, first mover, contracts, ecosystem capture). Note that the last three are business moats, not legal ones — a team with no patentable technology can still score well by owning a distribution relationship or a dataset.

Practicality does not require a prototype — and this is under-exploited

The Chief Judge is unambiguous:

"Teams can attain 5 points for Practicality without a prototype."

The cited example: a jet-engine electrolyser hydrogen burner — impossible to prototype on a student budget — earned 5 on Practicality and 4–5 on Innovation via component-technology explanation, credible drawings, technical explanation, cited expert testimony, and exploratory talks with manufacturers.

The evidence ladder (any one or more establishes proof of concept):

  1. Existing applications of the component technologies
  2. Expert testimony — a named person with credentials who reviewed it
  3. Research verifying feasibility — cited literature
  4. Convincing graphic representation — CAD, schematic, animation
  5. Partial or full prototype or demonstration
  6. Describing further research or experiments that would verify feasibility

Rungs 1–4 and 6 cost no money. Rung 2 costs one polite email. Most teams skip straight to "we couldn't afford a prototype" and leave 20% of the score on the table.

Finances: the cheapest points in the competition

Judges are told to be lenient here and it is only 10%. The Chief Judge's own scale:

Score What earns it
2–3 A reasonable effort; financials that add up to totals and show expected expense categories
4 Sound internal logic — even if actual market costs weren't surveyed
5 Strong logic, strong pricing strategy, and unit profit or supporting research/comparison

"A terrific team and product should not be knocked out of final consideration solely due to financial projections — as long as they made an effort with some logic involved."

Translation: 4/5 on Finances is available to any team that computes one unit's cost, one unit's price, and shows the arithmetic. Ocean Energy got there. Greencrete's "$5–10 thousand" did not.

The classic student error the Chief Judge names: costing only the immediate things (prototype materials) and never the full development and deployment.

Storytelling is where the video and website actually count

Storytelling & Professionalism (20%) explicitly "encompasses questions 1, 2 + video + entire submission and all attachments." Its sub-criteria include Credibility Boost (do the video and website reinforce expertise?) and Polish & Consistency.

This is why Greencrete's Canva /edit link and Tinkercad tree are not cosmetic issues — they are scored. And it is why judges are told to correct spelling in their own comments: the competition takes professional presentation seriously in both directions.

Concretely, judges reacted badly to casual video narration:

"'Yeah, this is like three-stage, this is like seven-stage or something, I don't know, I didn't count it' significantly lowers my belief that you'll champion this product and use investor resources well."

What judges say about exaggerated claims

If a team claims sales or deployments that seem inflated, judges are told not to award points for business achievements at all — the rubric doesn't reward having made a sale — and to raise it dispassionately:

"5,000 units manufactured means a total cost of $2M; you also need $1M for your building complex; the Nike deal is worth $800,000. Do you mean these have already been achieved, and if so, where and how did you fund the $3M required investment? Or are these projections for the future? I am uncertain."

Implication for teams: never inflate. It converts a neutral section into an active credibility problem, and it is trivially detected by arithmetic.


Part IV — The improvement ladder

If you are a team (or a coach) with a draft, work these in order. Ordered by points-per-hour.

Tier 0 — Mechanical (2 hours, worth ~15 points at the low end)

  1. Every question answers the question asked. Re-read the prompt, then your answer.
  2. No duplicated text between questions.
  3. Spellcheck. Then read aloud.
  4. Website is published and opens in an incognito window. Video link plays without login.
  5. References PDF exists, is complete, and includes any AI tools used.

Tier 1 — The originality search (3 hours, decides your 30% block) 6. Search for your innovation as a product category, not as your specific design. Search in the language of the industry, not the language of students. 7. Find the three strongest existing solutions — including commercial products, not just papers. 8. Name them in Q7. State honestly what they do better. 9. Rewrite Q4 to defend the delta — what specifically is new, and why can't the incumbent just do it? 10. Answer IP Defensibility with a real mechanism: patent claim, trade secret, dataset, first-mover contract, ecosystem lock-in. Not "we'll copyright our code."

Tier 2 — Mechanism review (1 expert-hour, prevents the ~75 ceiling) 11. Find one adult with domain expertise — not your coach, not a business person — and ask them one question: "Is there a reason this can't work?" 12. Specifically audit energy and mass balance. Where does the energy come from? Is anything a closed loop? Does conservation hold? 13. If there's a flaw, reframe rather than abandon. The honest constrained version usually scores higher than the ambitious impossible one.

Tier 3 — Evidence (5 hours, worth up to 20%) 14. Climb the evidence ladder as far as budget allows. Aim for ≥3 rungs. 15. If you have a prototype, measure something and plot it. 16. Email 2–3 real experts or potential customers. Quote them with names and titles in Q5. 17. State your prototype's limitations explicitly.

Tier 4 — Business arithmetic (4 hours, worth ~10–20 points) 18. One unit: bill of materials → unit cost → price → unit margin. Show the table. 19. Full development cost to market, not just prototype cost. 20. Market size with a source and a date. 21. Segment customers and identify where buyer ≠ payer. 22. Use of funds as a percentage split.

Tier 5 — Narrative (3 hours, worth up to 20%) 23. Q1 Elevator Pitch: problem → mechanism → who buys → why now. 150 words, no adjectives. 24. Q2 Team: roles and relevant capability only. Delete hobbies. 25. Video: rehearse. No "I don't know." Show the model doing something. 26. Website: consistent brand, working links, the same model image as the video.


Part V — Self-scoring rubric card

Score yourself 1–5 per theme, honestly. Multiply: Innovation×6 + Storytelling×4 + Practicality×4 + Marketing×4 + Finances×2 = /100.

Innovation — Can I name the 3 strongest existing solutions and state my defensible delta? Would a stranger's 10-minute Google search find something that makes me look uninformed? Do I have a real moat?

Storytelling — Would an investor read past Q1? Are the video, website, and brief telling one consistent story? Is there a single typo, broken link, or unpublished page?

Practicality — How many rungs of the evidence ladder do I occupy? Has a domain expert confirmed no mechanism-level flaw? Does energy/mass balance hold?

Marketing — Have I named real competitors and real customer segments? Do I know who pays versus who uses? Is my market size sourced?

Finances — Can I state the cost and price of exactly one unit, and the margin? Does my development budget cover the full path to market? Do the numbers add up?

Below 60: finish the submission (Tier 0–1). 60–75: you likely have a mechanism or originality problem (Tier 2). 75–85: you have a good project; evidence and unit economics are the gap (Tier 3–4). 85+: finalist territory. Now rehearse the pitch.

评审手册 —— 你的提交物到底经历了什么

来源:2024-25《Innovation Stage Judge Guide》、《Scoring Guide》、首席评委 Simon Glinsky 的 Q&A, 以及对 2025 年我被分配评审的五份真实 Innovation Brief 的精读 (Greencrete、SUNSYNC、Ocean Energy Dynamics、P-Bump、Puppy WC —— 全部为 Energy & Environment 赛道,全球组)。

本文只回答一个问题:2 分和 4 分之间差的是什么?


第一部分 —— 评委的真实处境

在动笔写任何东西之前,先把这个内化。

你的评委是一位志愿的专业人士 —— 创业者、风投、工程师、科学家或教育者。他们的约束条件:

  • 手上有 5 支队伍,总共 5–10 小时。也就是留给你约 90 分钟,还要写完五段散文体意见。
  • 他们被明确要求去 Google 你的说法:"执行在线搜索,以核实该方法或创新的原创性。"
  • 被要求先读完整份提交再打分,并且要分多次、隔天重读。
  • 看视频、打开网站
  • 打开你的参考文献 PDF
  • 各维度独立打分 —— 一个机制有问题的创新,仍然可以在市场维度拿 4 分。
  • 永远不会联系你。 他们心中的每一个疑问都得不到解答 —— 然后变成怀疑。

五个评委问题,按被问到的顺序

  1. (创新) 我以前见过这个吗?→ 他们会搜索
  2. (可行性) 这真的能行吗?→ 他们会用物理/领域常识去检验
  3. (叙事) 我相信这些人吗?→ 他们会看打磨程度和内在一致性
  4. (市场) 他们知道谁会买吗?→ 他们会找具名的客户和具名的竞品
  5. (财务) 数字对得上吗?→ 他们会算,但很宽容

最重要的一个推论: 因为评委无法向你提问,每一处含糊都会被判给你不利的一边。 "我不禁在想……"(I am left wondering)就是那句要扣分的话。


第二部分 —— 五份简述的评分与诊断

我是作为受指派的评委读它们的。名称按原样保留。下列分数是我按官方评分表做出的判断, 不是 Conrad 的官方分数。

校准表

简述 创新(30) 叙事(20) 可行(20) 市场(20) 财务(10) ~总分 结论
Ocean Energy Dynamics 3 5 2 5 4 ~76 极出色的商业写作,物理有问题
P-Bump 3 4 2 4 3 ~66 真工程,能量核算致命错误
SUNSYNC 2 3 4 3 3 ~58 真原型,创新不原创
Puppy WC 1 3 2 3 2 ~42 文笔流畅,产品早已存在
Greencrete 1 1 1 1 1 ~20 没做完

没有一份能进决赛。决赛需要 ~85–90+。请注意什么与总分无关:字数、文笔流畅度、热情。


案例一 —— Greencrete(~20/100):1 分是怎么炼成的

用仿生学做低碳混凝土。两名学生。

评委一眼就能看到的:

  • 第 2 题(团队)在讲滑雪和打篮球的爱好,从头到尾没说任何一个分工。150 词全部浪费。
  • 第 7 题(竞争)答非所问 —— 重复了目标客户的内容,没有列出任何竞争对手。
  • 第 8 题(进入市场)从第 6 题逐字复制粘贴
  • 第 9 题说商业模式是"listening(倾听)"—— 是 licensing(授权)的错别字,没改。
  • 第 5 题(验证)开头就写:"我们其实并没有一个 100% 能用的完整方案。"
  • 第 10 题(融资):把一种新水泥化学体系商业化的总开发成本是"5–10 千美元"。
  • 创新配图:一个棕色盒子上的 Tinkercad 圣诞树
  • 网站:一个原始的 Canva /edit 链接 —— 评委点进去可能进的是编辑器,不是网站。

诊断:这不是一个坏想法,这是一份没做完的提交。 低碳混凝土是真实的、数十亿美元级的问题; Holcim 也被正确点名了。这个队伍原本抓到了一条真线索(仿生学 → 螺旋纤维增强),然后放弃了。

首席评委在这种情况下的指示: 诚实打分 —— 全线 1 和 2 —— 但要像教练一样写意见。 "这支队伍拿一两分是合理的……这满足了目标 #2。我们的第一目标是为学生创造学习,包括那些表现不佳的。"

教学价值:五份中最高。 每一个失败都是机械性的、一个下午就能修好的: 答所问的那个问题、别复制粘贴、校对、把网站正确发布。


案例二 —— Puppy WC(~42/100):流畅的文笔救不了"不是创新"

带传感器的自清洁、可冲水宠物马桶。两名学生,韩国。

文笔实际上是五份里第二好的:段落干净,还有一套系统的竞品对比, 对上了尿垫、可洗垫和人造草坪垫。

而它的创新维度只有 1 分,因为评委被要求做原创性搜索 —— 而自动自清洁冲水宠物马桶是一个已经存在的消费品品类(Inubox、BrilliantPad 等)。 这份简述从头到尾没有提到它们存在。它只把自己定位在尿垫和垫子的对立面 —— 而那正是能让产品显得新颖的那个对比集。

雪上加霜的:

  • 全文没有任何工程细节 —— 没有传感器型号、没有每次冲水量、没有功耗、没有成本。
  • "专有的冲水机构……为宠物排泄物的特定稠度而设计" —— 断言了,从未描述。
  • 量化是一个没有依据的区间:"每年数百到一千多片一次性尿垫"。

这条教训,也是本文最重要的一条:

挑一个对自己有利的对比集,是队伍摧毁自己创新分数最常见的方式。 对比集由评委决定,不是你决定。 如果一个竞品存在而你没点名, 评委会得出结论:你要么没查,要么在藏。这两者都比那个竞品本身更糟。

解法不是换一个产品 —— 而是在第 7 题里点名真正的在位者,然后诚实地把差异化挣回来 ("现有产品售价 500 美元以上且需要专有耗材盒,而我们……")。


案例三 —— SUNSYNC(~58/100):一个能跑的原型,但里面没有新东西

会追踪太阳的太阳能板(光敏电阻 + 舵机 + Arduino),并用凸透镜聚光。 三名学生,City of Knowledge Academy,尼日利亚。

这支队伍做的真实工作比这组里任何一支都多,却排在中游。这就是教训所在。

评委真心认可的优点:

  • 一个做出来、测过的原型。可行性 = 4。
  • 真实的验证描述:他们写了如何遮住一个光敏电阻来验证舵机响应, 以及如何核对 9V 太阳能板与 Arduino 的 5V/3.3V 逻辑电平以免烧板。 这是真实的工程过程,读起来也确实真实。
  • 结构清晰,有引用来源,分工明确,用 SDG 7 框定。

创新为什么只有 2 分:

  • 太阳跟踪是 1960 年代的技术;聚光光伏是成熟领域。这两者都是评委被要求做的那个原创性搜索的第一批结果。
  • 这个组合并未被论证为新颖 —— 没有任何论证说明"跟踪器 + 透镜"合在一起大于两者之和。
  • "我们用了一种独特的方法给 Arduino 编程,所以我们的代码无法被复制。" 这一句是主动扣分的。对一位工程师评委而言,它暴露了对软件、对逆向工程、 对什么是知识产权的根本性误解。IP 可防御性是一条明确的子标准;这句话自信地答错了它。
  • "发电量是普通太阳能板的两倍,最高可达 90%" —— 一句话里两个互不相容的说法,没有任何测量, 而这支队伍手上就有一台能测的装置。

财务:120 美元原型 → 1,200 美元商用单元,没有任何解释这 10 倍差距的桥梁。 铺开阶段的"初期 500 美元"根本不是一份铺开预算。

诊断:他们手上就有回答问题的仪器,却没有做那个实验。 只需要一个下午,记录"跟踪+聚光"与"固定"两块板的瓦时数据并画成图, 就能同时抬高创新和可行性 —— 并把一句营销话术变成证据。

SUNSYNC 法则:做出来了就去测量它。 一个没被测量过的原型, 价值低于一份论证扎实的设计,因为它证明了你有条件去看却选择不看。


案例四 —— P-Bump(~66/100):算术严谨,物理错了

压电减速带:液压 10:1 力放大到 PZT 压电堆上,面向印度尼西亚。三名学生。

这组里技术上最有野心的一份。 真实的规格: 10 cm → 31.6 cm 缸径(面积比确实是 10:1)、4 段各 100 层的 PZT 堆、28×28 cm 底面、 真空回位机构、指向 ScienceDirect 和交通管理部门数据的行内 URL、 一个具名的真实竞品(涩谷站的压电地板)并诚实列出其优势劣势、 实用新型专利 + 商业秘密的组合策略,以及一个包含购电协议(PPA)的、逻辑自洽的 B2G 模式。

这在结构上正是市场和可行性拿 4–5 分该有的样子。

而可行性依然只有 2 分,原因只有一个,这个领域的评委 30 秒就能抓到:

减速带不会采集免费能量。它采集的是车辆发动机的能量。

压下减速带是对车辆做功 —— 额外的滚动阻力、额外烧掉的燃油, 先经过内燃机效率(约 25%),再经过压电转换。这个装置是一台净耗能机器, 悄悄向每一位司机征税。简述里对此只字未提。所有可信的压电道路试点,都恰恰死在这个经济性上。

然后是那些看起来严谨、其实不然的算术:

  • 声称 PZT 堆上的应力:46,875 Pa,即约 0.05 MPa。PZT 通常工作在数十 MPa。 这个几何结构把力放大的效果自己抵消掉了约一千倍。
  • 每辆车的能量用 E = P·V、1 mm 位移算出 4,113 J,再乘以每天 2 万辆车, 得出"可以供一户家庭用电"。一个 28×28 cm 的陶瓷堆在那个应力下根本不可能有 1 mm 位移, 而这个公式还把压力-体积功与材料 88% "效率"下的可回收电能混为一谈。

这条教训与 Puppy WC 那条正好相反:

数字不产生可信度。数字产生可被核查的主张。 一个具体的错误数字比一个诚实的区间更糟,因为它可证伪,而评委被要求去证伪它。

补救的路一直都在,而且很便宜:加一段现实性检验的说明。 "能量取自车辆,因此本方案只在减速带本来就是必需的交通稳静化场景下成立, 此时边际油耗被接受为该安全设施的代价。" 这个重构诚实、经得起推敲,而且更有意思


案例五 —— Ocean Energy Dynamics(~76/100):最好的商业写作,最糟的物理

集成进船体的波浪能系统:单向进水口、防杂物滤网、降湍流栅格、滑杆上的铅制振子、 密封液压缸、碳纤维液压涡轮。四名学生。

这就是决赛级的市场、财务和叙事写作长什么样。请研究这几点:

  • 点名真实竞品:CalWave、Mocean Energy、Carnegie Clean Energy —— 并给出精确的定位缺口 ("固定式系统……完全不适用于移动中的船只")。
  • 对客户做了细分,并区分了买单方与使用方:游艇主(主要)、造船厂(次要)、 租赁公司(第三层)、国防(相邻)—— 每一类都有自己的购买理由。
  • 给市场定了规模并给了轨迹:造船业 2,400 亿美元(2023)→ 4,600 亿美元(2033)。
  • 真实的分级单价:小型游艇单机 5 万美元;超级游艇多机 10 万–25 万美元。
  • 经常性收入:维护合同、授权、高级监控服务。
  • 资金用途按百分比拆分:1,000 万美元 = 20% 研发 / 40% 制造 / 20% 试点 / 20% 市场。
  • 分层 IP 策略:栅格与液压回路申请专利,制造工艺与材料配方作为商业秘密,品牌注册商标。
  • 诚实说出劣势:初始成本更高,然后用细分逻辑回应(富裕、注重形象的买家)。
  • 附件用得很对:一个 3D 模型、一段渲染动画,以及一个局部实物原型 (涡轮 + 直流电机 + Arduino),并且把限制直白写出来 ("我们无法用水测试,因为电机不防水")。

最后那点值得强调:他们主动说出了原型做不到什么。 评委会因此加分。 这与 SUNSYNC 那个没测量过的 90% 声称正好相反。

那么为什么可行性只有 2 分、创新只有 3 分?

船在浪中的自身运动就是能量来源。从这个运动里取能会增加阻力, 提高船舶的推进需求。而简述接着说,回收的电力"驱动一套创新的推进系统" —— 这就把回路闭合成了永动机。

一位船舶工程背景的评委会立刻看到。这个想法一个站得住的版本 —— 在锚泊或漂航的船上为船电负载做寄生式取能,此时阻力代价无关紧要。 那个版本站得住、依然算有点新意,而且用同样的文笔可以把可行性拿到 4 分。

诊断:这支队伍的写作水平超过了他们的工程评审水平。 没有任何具备流体力学背景的成年人在提交前读过第 4 题。除此之外,一切都是决赛级的。


第三部分 —— 横向模式

四种失败原型

原型 症状 例子 代价
没做完 答案重复、错别字、答非所问、链接坏掉 Greencrete 全部
有利对比集 只对比弱替代品;真正的在位者未被点名 Puppy WC 创新 → 1
未测量的原型 东西做出来了,却从未量化;主张无支撑 SUNSYNC 创新、可行性
打磨盖过物理 出色的商业写作包裹着不成立的机制 Ocean Energy、P-Bump 可行性 → 2,总分封顶 ~75

最后一种最危险,因为它对队伍本身、对商科背景的指导老师、 对任何 AI 写作助手都是不可见的只有一位领域专家、带着看机制的眼光去读,才能抓到。

创新分数由一段话决定

创新占 30% —— 单一最大的一块。首席评委给出的判据:

"如果任何人都能抄走那个应用,它就不是一个强创新。"

以及那个改变一切的例子:一种新汽水口味得 1/5。同样这个新口味, 如果能减少蛀牙,就是 3–5 分 —— "一个有价值、可保护的创新,尽管它是既有产品的新应用。"

所以创新不是用技术复杂度衡量的,而是用可防御的差异化衡量的。 SUNSYNC 用的硬件比汽水那个例子先进得多,却只拿 2 分。

子标准把这一点写得很直白:Verification 验证(在线搜索是否排除了重复?)与 IP Defensibility 可防御性(专利、商业秘密、著作权、先发优势、合同、生态位锁定)。 注意后三项是商业护城河,不是法律护城河 —— 一支没有可申请专利技术的队伍, 完全可以靠拥有一条分销关系或一份数据集拿到好分数。

可行性不需要原型 —— 而这一点被严重低估

首席评委说得毫不含糊:

"队伍可以在没有原型的情况下拿到可行性满分 5 分。"

他举的例子:一个喷气发动机电解制氢燃烧器 —— 在学生预算下根本不可能做原型 —— 拿到了可行性 5 分、创新 4–5 分,靠的是元器件技术说明、可信的图纸、 技术阐述、引用的专家证言,以及与铭牌制造商的初步接洽。

证据阶梯(任意一项或多项即可构成概念验证):

  1. 元器件技术在别处的既有应用
  2. 专家证言 —— 一位有资历、审阅过它的具名人士
  3. 证实可行性的研究文献
  4. 有说服力的图形表达 —— CAD、原理图、动画
  5. 局部或完整原型或演示
  6. 描述将会验证可行性的进一步研究或实验

第 1–4 和第 6 级不花钱。第 2 级只需一封礼貌的邮件。 大多数队伍直接跳到"我们买不起原型",把 20% 的分数留在了桌上。

财务:全场最便宜的分数

评委被要求在这里宽容,而且它只占 10%。首席评委自己的标尺:

分数 靠什么拿到
2–3 做了合理的努力;财务数字能加总出总额,并列出了预期的费用类别
4 内在逻辑自洽 —— 即使没有调研实际市场成本
5 逻辑强、定价策略强、有单位利润或支撑性的研究/对比

"一支出色的队伍和产品,不应该仅仅因为财务预测就被挡在决赛之外 —— 只要他们做了努力且有一定逻辑。"

翻译过来:任何一支算清了单个单元的成本、单个单元的售价,并把算术写出来的队伍, 都能拿到财务 4/5 分。 Ocean Energy 做到了。Greencrete 的"5–10 千美元"没有。

首席评委点名的经典学生错误:只算眼前的成本(原型材料),从不算完整的开发与部署成本。

叙事:视频和网站真正计分的地方

叙事与专业度(20%)明确"涵盖第 1、2 题 + 视频 + 整份提交及所有附件"。 它的子标准包括 Credibility Boost 可信度加成(视频和网站是否强化了专业度?) 与 Polish & Consistency 打磨与一致性

这就是为什么 Greencrete 的 Canva /edit 链接和 Tinkercad 圣诞树不是外观问题 —— 它们计分。 这也是为什么评委被要求在自己的意见里改正拼写:这个比赛在两个方向上都认真对待专业呈现。

具体地,评委对随意的视频旁白反应很差:

"'嗯,这个大概是三级的,也可能是七级什么的,我不知道,我没数' 这句话 显著降低了我对你们能捍卫这个产品、善用投资人资源的信心。"

评委怎么看夸大的说法

如果一支队伍声称的销量或部署看起来虚高,评委被要求完全不为商业成就打分 —— 评分表本来就不奖励"做成过一笔生意" —— 并且要冷静地提出来:

"5,000 台的制造意味着总成本 200 万美元;你们还需要 100 万美元建厂区; 那个 Nike 合作价值 80 万美元。你们的意思是这些已经实现了吗?如果是, 那 300 万美元的必要投资是从哪里、怎么筹到的?还是说这些是未来的预测?我不确定。"

对队伍的含义:永远不要夸大。 它会把一个中性的章节变成一个主动的可信度问题, 而且用算术就能轻易识破。


第四部分 —— 改进阶梯

如果你是一支(或指导一支)已有草稿的队伍,按这个顺序做。按"每小时能拿多少分"排序。

第 0 层 —— 机械性(2 小时,在低分段值约 15 分)

  1. 每一题都回答被问的那个问题。先重读题干,再读你的答案。
  2. 题与题之间没有重复文本
  3. 拼写检查。然后朗读一遍。
  4. 网站已发布,能在无痕窗口打开。视频链接不登录就能播放。
  5. 参考文献 PDF 存在、完整,并包含用过的任何 AI 工具

第 1 层 —— 原创性搜索(3 小时,决定你 30% 的那一块) 6. 把你的创新当作一个产品品类去搜,而不是当作你的具体设计。 用行业的语言搜,不要用学生的语言。 7. 找出三个最强的现有解决方案 —— 包括商业产品,不只是论文。 8. 在第 7 题里点名它们。诚实说出它们哪里做得更好。 9. 重写第 4 题来捍卫这个差值 —— 具体新在哪,以及为什么在位者不能直接照做? 10. 用一个真实机制回答 IP 可防御性:专利权利要求、商业秘密、数据集、 先发合同、生态锁定。不要写"我们会给代码申请著作权"。

第 2 层 —— 机制评审(1 个专家小时,防止 ~75 分天花板) 11. 找一位有领域专长的成年人 —— 不是你的指导老师,不是商科的人。只问一个问题: "有没有什么原因让这个方案行不通?" 12. 特别审计能量与质量守恒。能量从哪来?有没有闭环?守恒成立吗? 13. 如果有缺陷,重构而不是放弃。诚实的受限版本通常比雄心勃勃的不可能版本得分更高。

第 3 层 —— 证据(5 小时,最高值 20%) 14. 在预算允许范围内尽量往证据阶梯上爬。目标 ≥3 级。 15. 如果你有原型,测点什么东西,并画成图。 16. 给 2–3 位真实专家或潜在客户发邮件。在第 5 题里带姓名和头衔引用他们。 17. 明确写出你的原型的局限

第 4 层 —— 商业算术(4 小时,值约 10–20 分) 18. 一个单元:物料清单 → 单位成本 → 售价 → 单位毛利。把表格放出来。 19. 到达市场的完整开发成本,不是原型成本。 20. 市场规模要有来源和日期。 21. 对客户做细分,并指出买单方 ≠ 使用方的地方。 22. 资金用途按百分比拆分。

第 5 层 —— 叙事(3 小时,最高值 20%) 23. 第 1 题电梯陈述:问题 → 机制 → 谁买 → 为什么是现在。150 词,不要形容词。 24. 第 2 题团队:只写分工和相关能力。 删掉爱好。 25. 视频:排练。不要说"我不知道"。让模型做点什么给人看。 26. 网站:品牌一致,链接可用,用与视频相同的模型图。


第五部分 —— 自评分卡

诚实地给自己每个维度打 1–5 分。计算: 创新×6 + 叙事×4 + 可行性×4 + 市场×4 + 财务×2 = /100。

创新 —— 我能不能点名三个最强的现有方案,并说出我可防御的差值? 一个陌生人 10 分钟的 Google 搜索,会不会找到让我显得很无知的东西?我有真的护城河吗?

叙事 —— 投资人会不会读过第 1 题继续往下看?视频、网站、简述讲的是同一个故事吗? 有没有一个错别字、坏链接或未发布的页面?

可行性 —— 我占了证据阶梯的几级?有领域专家确认过没有机制层面的缺陷吗? 能量/质量守恒成立吗?我做出来的东西,我测量了吗?

市场 —— 我点名了真实竞品和真实客户细分吗?我知道谁买单、谁使用吗? 我的市场规模有来源吗?

财务 —— 我能说出恰好一个单元的成本、售价和毛利吗? 我的开发预算覆盖了通往市场的完整路径吗?数字加得起来吗?

低于 60: 先把提交做完(第 0–1 层)。 60–75: 你多半有机制或原创性问题(第 2 层)。 75–85: 项目不错;差的是证据和单位经济(第 3–4 层)。 85+: 决赛区间。现在去练路演。