Dental Foundation Models & Benchmarks

Dental multimodal LLMs, clinical reasoning benchmarks, expert-calibrated dental evaluation, and OmniDentBench.

Dental LLMs, multimodal diagnosis, and expert-calibrated benchmarks
Dental Foundation Models DentalGPT GlobalDentBench OmniDentBench Clinical reasoning
Dental foundation model benchmark stack

牙科大模型是医疗 AI 里足够独立的一条垂直线:它需要理解口腔影像、牙科专科知识、病例推理、治疗风险、专家评分和跨地区诊疗标准。DentalGPT 提供多模态牙科模型路线,GlobalDentBench 和 OmniDentBench 则提供高难度、专家校准、临床推理导向的 benchmark 入口。

Research Storyline

Model
训练牙科多模态大模型

DentalGPT 用大规模牙科图像、专业 caption、instruction tuning 和 GRPO 强化多模态复杂推理。

Bench
构建跨国临床推理 benchmark

GlobalDentBench 覆盖 88 个国家/地区、14 个牙科专科和多种题型,强调专家校准和风险分析。

ODB
搭建开放评测平台

OmniDentBench 面向复杂临床决策和生物医学研究,提供全球牙科基准评测与 leaderboard 入口。

Context
对齐外部 dental benchmark

DentalBench、OralGPT-Omni 和 OralMLLM-Bench 等工作显示牙科领域正在形成专门的模型和评测生态。

Representative Work

Model
DentalGPT: Incentivizing Multimodal Complex Reasoning in Dentistry

Builds a specialized 7B dental MLLM with domain knowledge injection and reinforcement learning for dental visual reasoning.

Paper
Bench
GlobalDentBench

A multinational benchmark for LLM clinical reasoning in dentistry with expert calibration, specialty taxonomy, and risk analysis.

Paper
Platform
OmniDentBench

A global dental benchmarking platform for high-difficulty clinical decision-making and biomedical research evaluation.

Platform
QA
DentalBench

A bilingual dental QA benchmark and DentalCorpus for evaluating and adapting LLMs in dentistry.

Paper
MLLM
OralGPT-Omni

A dental-specialized multimodal LLM and MMOral-Uni benchmark for dental image analysis.

Paper

Benchmark Layers

Dental visual understanding

Intraoral photos, panoramic X-rays, periapical radiographs, and dental-specific VQA expose failures that generic medical MLLMs often miss.

Clinical reasoning complexity

GlobalDentBench moves from knowledge recall to routine reasoning and individualized reasoning with patient-specific constraints.

Expert calibration

Dentist-in-the-loop validation and rubric-based scoring make dental evaluation more clinically meaningful than generic QA accuracy.

Risk analysis

Dental AI needs to track unsafe treatment suggestions, irreversible harm risks, and specialty-specific failure modes.

Resource Map

OmniDentBench

Global dental benchmark platform and leaderboard.

Platform
DentalGPT

Dental multimodal model repository and evaluation setup.

Repository
GlobalDentBench

Expert-calibrated multinational dental clinical reasoning benchmark.

Repository