DNAfinity — Competitive Consensus (2026-07-11)

Code-truth rerun · median-of-4 independent model reviewers · rubric 7 มิติ × 0–5 · directional self-assessment (ยังไม่ผ่าน user research)

🟣 Claude Sonnet 4.6🟣 Claude Opus 4.8 — 17 read-only code subagents🟠 Codex🔵 Gemini

ทำไมรอบนี้ "ต่าง"

~3.7/5median consensus · 16 แอป
0 → shipHuman-DNA Engine deploy จริง 07-07 (moat ที่ 07-04 บอกเสี่ยงสุด)
14/16แอปที่ขยับขึ้น (รอบก่อนขยับ 2)
D7=1Traction เพดานทุกแอป — ยังไม่มี user ภายนอก
ยอมรับตรง ๆ: ตัวที่เข้มสุด (Opus 4.8) portfolio = 3.45 ส่วนหนึ่งเพราะรวมแอปใหม่ที่ให้คะแนนต่ำ — ถ้าปรับ roster ให้เท่ารอบก่อน portfolio แทบ flat (3.56 vs 3.59). ควรอ่านคู่กับพาดหัว "3.7 / ขึ้นทั้งกระดาน".

คะแนน consensus ต่อแอป (median-of-4)

AppSonnetOpus 4.8CodexGeminiMedianΔ 07-04Spread
WorkDNA4.434.34.44.44.40+0.10.13
SmeDNA4.434.04.14.44.25+0.150.43
BloomDNA4.143.94.14.14.10+0.20.24
AppForge4.003.94.04.04.000.00.10
MU-AI Core4.003.94.04.04.00+0.10.10
CareDNA4.143.73.94.03.95+0.250.44
FinDNA 4.003.43.93.93.90+0.20.60
MU-AI Game3.713.73.74.03.71+0.110.30
DataDNA3.713.43.63.73.65+0.250.31
EmotiQ3.713.33.43.73.55+0.250.41
TradeDNA3.573.33.43.63.49+0.190.30
PlantDNA new3.572.93.43.93.49new1.00
PetDNA3.432.93.43.63.42+0.320.70
MUAI-Verse 3.432.43.43.43.40+0.51.03
OperDNA3.293.33.43.33.300.00.11
InfluDNA3.293.03.33.33.300.00.30

Portfolio: Gemini 3.83 · Sonnet 3.74 · Codex 3.71 · Opus 4.8 3.45 → median 3.73 / mean 3.68. = contested (ดูด้านล่าง). Tier: Leader ≥4.0 · Contender 3.5–3.99 · Developing <3.5.

เห็นตรง vs เห็นต่าง

จุดแข็ง & ความเสี่ยง (consensus)

จุดแข็งความเสี่ยง
Human-DNA Engine ship + wire ข้ามแอป — moat ที่ copy ยากTraction=1 ทุกแอป — ยังไม่มี user ภายนอก (blocker #1)
Ecosystem + ไทย compliance-as-code (คู่แข่ง ~2 ทั้งสองมิติ)งานเด่นหลายชิ้นยังอยู่ branch ไม่ deploy
วินัย honesty — ทุกแอปมี claim-matrix บล็อกการเคลมเกินจริงของตัวเอง"Ecosystem" บางที่เป็นแค่ SSO ไม่ใช่ DNA wiring จริง
First-value + agentic pattern แพร่เร็วทั้ง estateAgentic gap — "สมองครบ มือยังไม่ขยับ" (คู่แข่ง execute แล้ว)
Founder+AI leverage ~45.3× (332.58 vs 7.339 FTE-month)ขาด integration ภายนอก (bank/broker/marketplace/device)

Pull-quotes

"Traction remains the portfolio bottleneck. Every active app stays D7=1 … branch evidence counts for product readiness, not market trust."— Codex (verbatim)
"Human-DNA went from architecture on paper to a deployed engine with the first producer running — in 5 days — closing the gap the 07-04 report called the highest risk."— Claude Sonnet 4.6
"The largest systemic upgrade since development began … a real ecosystem moat, not just a claim."— Gemini
"Brain complete, hands don't move — most apps top out at D2=4 without a scheduler, while Cleo and Lindy already execute in production."— Claude Opus 4.8

Consolidated 2026-07-11 by the DNAfinity portal from 4 independent model reviews (median-of-4). Directional self-assessment for internal product prioritization — not user-research-validated. Source: docs/competitive/CONSENSUS-2026-07-11.md + per-model *-2026-07-11 reports.