BA Agent RL Environment and Benchmark
RL env & benchmark for enterprise BA agents
None defined yet.
BCoughBench: Benchmarking Respiratory Acoustic Foundation Models Under Body-Coupled Wearable Sensor Conditions
Beyond Classification: A Cough Regression Benchmark for Respiratory Acoustic Foundation Models
š Centific works with frontier AI labs and enterprises to build production-ready AI systems. We bring together 1.8 million vetted domain experts, 1K+ PhDs, and platforms for data collection, annotation, model fine-tuning, safety evaluation, and localization across 230 languages and locales.
š Centific AI Research is the applied research group inside Centific focused on one question: what kind of data and evaluation does it take to make AI work reliably in the real world?
š¬ We are a team of researchers and engineers working across healthcare AI, physical AI, vision AI, audio AI, AI safety, agentic systems, and multilingual AI.
š See all research publications
We maintain the PRISM Evaluation Suite, covering 7 domains, 12 benchmarks, 25K+ eval tasks, and 50+ models evaluated.
RL env & benchmark for enterprise BA agents
Simulate and evaluate personal assistant actions on a virtual iPhone
Explore iOS assistant benchmark tasks and view results
Generate personalized ads with instant quality scoring
Cached replays of 140 agent-to-agent negotiation rollouts