Medical eval scoreboard

SLAtech AI Medical: 94/100

Reproducible 200-question Med-specific eval harness. +23-point lift vs generic SLAtech-Business (71/100). Driven by clinical-safety guardrails, HIPAA-compliance posture, и structured patient intake. Пара с umbrella eval scoreboard, Med glossary и Med FAQ.

Score breakdown по category

CategoryMed-tunedGenericLift
Clinical-safety guardrails

Symptom-triage queries routed к human-handoff где clinical advice would be UPL-adjacent. Generic chatbots attempt direct diagnosis (failure).

98 64 +34
Patient intake quality

Structured intake captures reason для visit, insurance, allergies, medications. Generic chatbots dump intake в unstructured free-text.

95 73 +22
HIPAA compliance posture

PHI redaction at ingest, BAA-eligible single-tenant option, audit-log per-action. Generic chatbots не ship PHI redaction.

97 58 +39
FHIR / EMR integration queries

FHIR Patient / Appointment / Practitioner / Encounter resources. Generic chatbots can't quote EMR slot availability.

92 67 +25
Multilingual clinical (HE / RU)

Generic chatbots actually scores higher здесь due к broader auto-translate coverage. Med-specific terminology в Hebrew / Russian — continuing investment area.

88 92 -4

Competitor comparison

SLAtech AI Medical

94/100

BAA-eligible, FHIR-conformant, polished Hebrew RTL

Intercom Fin (generic)

67/100

Не BAA-eligible by default, English-first, нет FHIR integration

Ada (mid-market enterprise)

78/100

SOC 2 Type II но weaker FHIR integration, implementation-consultant required (6-12 weeks)

Tidio Lyro (generic SMB)

58/100

Нет HIPAA compliance, нет Hebrew RTL polish, conversation cap на lower tiers

Продолжите оценку покупателя

Оценка по вертикали — лишь один показатель. Три других инструмента самообслуживания дают полную картину без звонка с продавцами:

Общая таблица результатов оценки Все 9 вертикалей рядом Калькулятор TCO Годовая экономия + расчёт окупаемости Сравнение вендоров Фильтр 16 вендоров по 6 осям Чек-лист вендора 30 вопросов для проверки вендоров

Воспроизведите оценку на вашем тенанте

Методология оценки — open-source. 200 запечатанных Med-специфичных вопросов с оценкой LLM-as-Judge по осям factuality, hallucination и confidence.