Voice AI in production
Engineering a healthcare voice agent with real-time scheduling integrations.
I’m Alex Chandler, an AI Scientist at BCG X. I work on AI evaluation, applied research, and software engineering. This site brings together some of the systems I’ve built, the research I’ve published, and projects I’ve pursued along the way.

Explore 22 projects, publications and engineering contributions, from current work to the earliest experiments.
Engineering a healthcare voice agent with real-time scheduling integrations.
A simulation harness for realistic caller personas, automated scoring and behavioral regression tests.
Systems and methods for measuring performance of large language models.
Assessment, verification and correction workflows for generated reports.
Backend and DevOps engineering for an agentic AI application deployed for government use, serving 40K+ users.
LLM Efficiency Research
Fine-tuned language models that answer batches of questions in one inference pass.
AI Evaluation Research
Instruction-tuning for financial claim classification and reasoned explanations.
AI Evaluation Research
Token-level detection of hallucinated spans using features from internal LLM layers.
Reinforcement Learning Research
Comparing reinforcement learning and counterfactual regret minimization in imperfect-information poker.
AI Evaluation Research
Ensembling prompt judgments to detect factual inconsistencies in generated summaries.
AI Evaluation Research
Context decomposition and prompt-based evaluation of factual consistency.
Distributed Computing Research
Integrating stable membership views into a consensus system.
LLM Efficiency Research
Give a language model one document and several questions or requested outputs. Return the answers together in one inference pass.
Engineering a financial chatbot with named-entity recognition, table extraction and automated answer evaluation.
Computer Vision Research
Contrastive learning, hard negatives and temporal reordering for video-language representations.
Reinforcement Learning Research
Reformulating robot imitation learning as a sequence of returns, states, and actions.
Language Model Research
Topical control through contrasting expert language models and self-organizing structures.
APIs, test automation and AI threat detection for network visibility products.
AI Safety Evaluation
Testing question-answering robustness by using modified beam search to find adversarial word sequences that reduce ELECTRA's accuracy.
Applied Computer Vision
Computer vision models for instrument identification and valuation, supported by a database of 10,000+ instruments.
Reinforcement Learning Research
Planning well paths through 3D subsurface data with actor-critic methods.
Geospatial Analytics & Software Engineering
Analyzing political districts with graph theory and computational geometry; leading backend, API development and data collection for the GerryMap web application.
Voice AI, evaluation, and enterprise systems.
Explore work ↗Jul 2024–Jul 2025Enterprise AI applications, backend engineering and DevOps.
Explore work ↗Jun–Aug 2023Financial question answering, document extraction and automated evaluation.
Explore work ↗Oct 2020–Mar 2022Software engineering for network visibility products.
Explore work ↗2019–2024A foundation in computer science, artificial intelligence and machine learning.
Explore work ↗Building AI applications, evaluating their behavior,
and researching the methods behind them.