About
I am a Research Scientist in Artificial Intelligence Benchmarking and Evaluation at IBM Research.
I work at the intersection of Natural Language Processing and Machine Learning, under the guidance of Michal Shmueli-Scheuer,
and in close collaboration with Leshem Choshen
and Ariel Gera.
My research aims to develop benchmarks and methodologies that go beyond basic evaluation to map the boundaries of AI systems and expose
their abilities to help us understand them better. This includes exploring their failure modes, robustness, and behavior under stress.
I've also recently started studying physics, and I am excited about combining AI with other sciences to drive new breakthroughs.
I completed my MSc in Computer Science specializing in NLP at Bar-Ilan University, where I was advised by Prof. Yoav Goldberg.
I am always open to new collaborations and love brainstorming about AI, benchmarking, or scientific breakthroughs - please feel free to reach out!
Research Interests
Selected Publications
ErrorMap and ErrorAtlas: Charting the Failure Landscape of Large Language Models
arXiv:2601.15812