Voice AI in production
Engineering a healthcare voice agent with real-time scheduling integrations.
Research & engineering
Applied AI, backend engineering, computer vision, distributed systems and AI evaluation.
Through the years
Engineering a healthcare voice agent with real-time scheduling integrations.
A simulation harness for realistic caller personas, automated scoring and behavioral regression tests.
Systems and methods for measuring performance of large language models.
Assessment, verification and correction workflows for generated reports.
Backend and DevOps engineering for an agentic AI application deployed for government use, serving 40K+ users.
LLM Efficiency Research
Fine-tuned language models that answer batches of questions in one inference pass.
AI Evaluation Research
Instruction-tuning for financial claim classification and reasoned explanations.
AI Evaluation Research
Token-level detection of hallucinated spans using features from internal LLM layers.
Reinforcement Learning Research
Comparing reinforcement learning and counterfactual regret minimization in imperfect-information poker.
AI Evaluation Research
Ensembling prompt judgments to detect factual inconsistencies in generated summaries.
AI Evaluation Research
Context decomposition and prompt-based evaluation of factual consistency.
Distributed Computing Research
Integrating stable membership views into a consensus system.
LLM Efficiency Research
Give a language model one document and several questions or requested outputs. Return the answers together in one inference pass.
Engineering a financial chatbot with named-entity recognition, table extraction and automated answer evaluation.
Computer Vision Research
Contrastive learning, hard negatives and temporal reordering for video-language representations.
Reinforcement Learning Research
Reformulating robot imitation learning as a sequence of returns, states, and actions.
Language Model Research
Topical control through contrasting expert language models and self-organizing structures.
APIs, test automation and AI threat detection for network visibility products.
AI Safety Evaluation
Testing question-answering robustness by using modified beam search to find adversarial word sequences that reduce ELECTRA's accuracy.
Applied Computer Vision
Computer vision models for instrument identification and valuation, supported by a database of 10,000+ instruments.
Reinforcement Learning Research
Planning well paths through 3D subsurface data with actor-critic methods.
Geospatial Analytics & Software Engineering
Analyzing political districts with graph theory and computational geometry; leading backend, API development and data collection for the GerryMap web application.
No work matches these filters. Clear filters to see all projects.