Abel Yagubyan
Senior Data Scientist at C3.ai
Hello! I'm a Senior Data Scientist at C3.ai with expertise in Deep Learning, Predictive Maintenance, and High-Performance Computing. I hold an M.S. in Computer Science from Northwestern University (Summa Cum Laude) and dual B.A. degrees in Computer Science & Applied Mathematics from UC Berkeley.
Previously, I co-founded FibonAI (UC Berkeley Skydeck), conducted research at Lawrence Berkeley National Laboratory on UPC++ performance testing, and interned at Apple. I've also reviewed 100+ papers for Elsevier's AI and Vision Computing journals.
News
Named Area Triager on TruLens (the open-source LLM and agent evaluation library maintained by Snowflake) and added to MAINTAINERS.md.
Contribution credited in inspect_evals v0.20.0, the UK AI Security Institute's evaluation suite for Inspect AI.
"The Coin Flip Judge?" accepted at GroundLM 2026 (EMNLP 2026) as an Archival Long Paper; camera-ready submitted.
Work cited by research groups at Meta, Alibaba and Nanjing University, USTC, University of New Mexico, and KTH.
Released two sole-author preprints on LLM agent reproducibility and LLM-as-a-judge reliability.
Promoted to Senior Data Scientist at C3.ai in under 18 months for high-impact customer solutions
Co-founded FibonAI, accepted into UC Berkeley's Skydeck Pad-13 Incubator competing against 5,000+ startups
Graduated with M.S. in Computer Science from Northwestern University with Summa Cum Laude Honors
Research contributor at Lawrence Berkeley National Lab's Pagoda Project on UPC++ performance testing
Published "Embedding of Programming IDEs into Computer-Based Testing Software" at ACM SIGCSE '22
Graduated from UC Berkeley with dual B.A. in Computer Science & Applied Mathematics
Software Engineering Intern at Apple x UC Berkeley, managing CS61C projects for 4000+ HBCU students
Research & Publications
How Consistent Are LLM Agents? Measuring Behavioral Reproducibility in Multi-Step Tool-Calling Pipelines
Embedding of Programming IDEs into Computer-Based Testing Software
SCALPEL: Customized Deep Neural Network Compression
UPC++ Performance Regression Testing & Benchmarking
Citations & Adoption
Cited and built on by research groups at Meta Superintelligence Labs (OmnilingualGAIA2, arXiv:2608.08775), Alibaba Group and Nanjing University (OpenCodeReview, arXiv:2608.09290, open source at github.com/alibaba/open-code-review), USTC (arXiv:2607.02873), University of New Mexico, and KTH Royal Institute of Technology (arXiv:2609.06147). 25 citations total as of September 2026.
Waxell published operational guidance for its users based on the judge-reliability findings, citing the paper as its first source.
Industry Experience
C3.ai
FibonAI
Apple x UC Berkeley
Software & Projects
Open Source
TruLens
inspect_evals
Awards & Recognition
Peer Reviewer
Engineering Applications of Artificial Intelligence, Image and Vision Computing, Neural Networks, Information Fusion, Neurocomputing, Expert Systems with Applications, Knowledge-Based Systems, and Data and Knowledge Engineering.
Elsevier Recognised Reviewer Certificates
Certificates awarded for review volume across the following journals:
GroundLM 2026 Workshop at EMNLP 2026
"The Coin Flip Judge?" accepted as an Archival Long Paper (Program Chairs decision: Accept; two area-chair recommendations to accept).
UC Berkeley Skydeck Pad-13
FibonAI accepted into UC Berkeley's Skydeck Pad-13 Incubator (competing against 5,000+ startups globally).