11:00
40d ago
→GPT-5.5 Instant brings frontier-level health responses to free ChatGPT users
OpenAI claims GPT-5.5 Instant now matches frontier models on health benchmarks like HealthBench Professional, and physician panels rated its responses higher than human-written ones on accuracy and communication. Production monitoring shows a 71% drop in flagged factual errors over two months. The improvements come from model advances plus a global physician network that defines evaluation rubrics and failure modes. The post doesn't disclose absolute error rates or detailed blind-review scores.
72
SCORE
H1·K1·R0