Healthcare Document-Grounded QA Benchmark (Extending GDP.pdf)
Document-work benchmark for healthcare persona
None defined yet.
BCoughBench: Benchmarking Respiratory Acoustic Foundation Models Under Body-Coupled Wearable Sensor Conditions
Beyond Classification: A Cough Regression Benchmark for Respiratory Acoustic Foundation Models
Centific works with frontier AI labs and enterprises to build production-ready AI systems. We bring together 1.8 million vetted domain experts, 1K+ PhDs, and platforms for data collection, annotation, model fine-tuning, safety evaluation, and localization across 230 languages and locales.
Centific AI Research is the applied research group inside Centific focused on one question: what kind of data and evaluation does it take to make AI work reliably in the real world?
We are a team of researchers and engineers working across healthcare AI, physical AI, vision AI, audio AI, AI safety, agentic systems, and multilingual AI.
š See all research publications
We maintain the PRISM Evaluation Suite, covering 7 domains, 12 benchmarks, 25K+ eval tasks, and 50+ models evaluated.
Document-work benchmark for healthcare persona
RL env & benchmark for enterprise BA agents
Simulate and evaluate personal assistant actions on a virtual iPhone
Explore iOS assistant benchmark tasks and view their run details