default search action
Eliya Habba
Person information
Refine list
refinements active!
zoomed in on ?? of ?? records
view refined list in
2020 – today
- 2026
- [c7]Shahar Levy, Eliya Habba, Reshef Mintz, Barak Raveh, Renana Keydar, Gabriel Stanovsky:
ScheMatiQ: From Research Question to Structured Data through Interactive Schema Discovery. ACL (3) 2026: 220-230 - [i14]Mubashara Akhtar, Anka Reuel, Prajna Soni, Sanchit Ahuja, Pawan Sasanka Ammanamanchi, Ruchit Rawal, Vilém Zouhar, Srishti Yadav, Chenxi Whitehouse, Dayeon Ki, Jennifer Mickel, Leshem Choshen
, Marek Suppa, Jan Batzner, Jenny Chim, Jeba Sania, Yanan Long, Hossein A. Rahmani, Christina Knight, Yiyang Nan, Jyoutir Raj, Yu Fan, Shubham Singh, Subramanyam Sahoo, Eliya Habba, Usman Gohar, Siddhesh Pawar, Robert Scholz, Arjun Subramonian, Jingwei Ni, Mykel J. Kochenderfer, Sanmi Koyejo, Mrinmaya Sachan, Stella Biderman, Zeerak Talat, Avijit Ghosh, Irene Solaiman:
When AI Benchmarks Plateau: A Systematic Study of Benchmark Saturation. CoRR abs/2602.16763 (2026) - [i13]Shahar Levy, Eliya Habba, Reshef Mintz, Barak Raveh, Renana Keydar, Gabriel Stanovsky:
ScheMatiQ: From Research Question to Structured Data through Interactive Schema Discovery. CoRR abs/2604.09237 (2026) - [i12]Eliya Habba, Itay Itzhak, Asaf Yehudai, Yotam Perlitz, Elron Bandel, Michal Shmueli-Scheuer, Leshem Choshen, Gabriel Stanovsky:
Growing Pains: Extensible and Efficient LLM Benchmarking Via Fixed Parameter Calibration. CoRR abs/2604.12843 (2026) - [i11]Itay Itzhak, Eliya Habba, Gabriel Stanovsky, Yonatan Belinkov:
From Feelings to Metrics: Understanding and Formalizing How Users Vibe-Test LLMs. CoRR abs/2604.14137 (2026) - [i10]Avijit Ghosh, Anka Reuel, Jenny Chim, Wm. Matthew Kennedy, Srishti Yadav, Jennifer Mickel, Yanan Long, Andrew Tran, Anastassia Kornilova, Damian Stachura, Kevin Klyman, Felix Friedrich, Jeba Sania, Jan Batzner, Anoop Mishra, Eliya Habba, Yixiong Hao, Nathan Heath, Shalaleh Rismani, Usman Gohar, Andrea Loehr, David Manheim, Ruchira Dhar, Sree Harsha Nelaturu, Aarush Sinha, Leshem Choshen, Drishti Sharma, Ishan Khire, Amit Saha, Subramanyam Sahoo, Michael Hardy, Michael Alexander Riegler, Kabir Manghnani, Michelle Lin, Yanan Jiang, Yilin Huang, Asaf Yehudai, Jessica Ji, Aris Hofmann, Mubashara Akhtar, Max Lamparth, Nuno Moniz, Yacine Jernite, Stella Biderman, Zeerak Talat, Sanmi Koyejo, Mykel J. Kochenderfer, Irene Solaiman:
Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting. CoRR abs/2606.09809 (2026) - [i9]Jan Batzner, Sree Harsha Nelaturu, Damian Stachura, Anastassia Kornilova, Jon Crall, Tommaso Cerruti, Yanan Long, Yifan Mai, Sanchit Ahuja, Asaf Yehudai, Marek Suppa, John P. Lalor, Oluwagbemike Olowe, Jatin Ganhotra, Brian H. Hu, Eliya Habba, Andrew M. Bean, Chang Liu, Sander Land, Steven Dillmann, Aniketh Garikaparthi, Elron Bandel, Saki Imai, James Edgell, Wm. Matthew Kennedy, Jenny Chim, Patrick Meusling, Asteria Kaeberlein, Venkata Ramachandra Karthik Chundi, Manasi Patwardhan, Martin Ku, Austin Meek, Leon Knauer, Brian Wingenroth
, Srishti Yadav, Usman Gohar, Felix Friedrich, Michelle Lin, Jennifer Mickel, Arman Cohan, Stella Biderman, Irene Solaiman, Zeerak Talat, Anka Reuel, Mubashara Akhtar, Gjergji Kasneci, Avijit Ghosh, Leshem Choshen
:
Every Eval Ever: A Unifying Schema and Community Repository for AI Evaluation Results. CoRR abs/2606.14516 (2026) - 2025
- [c6]Eliya Habba, Ofir Arviv, Itay Itzhak, Yotam Perlitz, Elron Bandel, Leshem Choshen
, Michal Shmueli-Scheuer, Gabriel Stanovsky:
DOVE: A Large-Scale Multi-Dimensional Predictions Dataset Towards Meaningful LLM Evaluation. ACL (Findings) 2025: 11744-11763 - [c5]Eliya Habba, Noam Dahan, Gili Lior, Gabriel Stanovsky:
PromptSuite: A Task-Agnostic Framework for Multi-Prompt Generation. EMNLP (System Demonstrations) 2025: 254-263 - [c4]Sarel Duanis, Asnat Greenstein-Messica, Eliya Habba:
JSON Whisperer: Efficient JSON Editing with LLMs. EMNLP (Industry Track) 2025: 1265-1274 - [c3]Gili Lior, Eliya Habba, Shahar Levy, Avi Caciularu, Gabriel Stanovsky:
ReliableEval: A Recipe for Stochastic LLM Evaluation via Method of Moments. EMNLP (Findings) 2025: 11146-11153 - [i8]Gabriel Stanovsky, Renana Keydar, Gadi Perl, Eliya Habba:
Beyond Benchmarks: On The False Promise of AI Regulation. CoRR abs/2501.15693 (2025) - [i7]Eliya Habba, Ofir Arviv, Itay Itzhak, Yotam Perlitz, Elron Bandel, Leshem Choshen, Michal Shmueli-Scheuer, Gabriel Stanovsky:
DOVE: A Large-Scale Multi-Dimensional Predictions Dataset Towards Meaningful LLM Evaluation. CoRR abs/2503.01622 (2025) - [i6]Gili Lior, Eliya Habba, Shahar Levy, Avi Caciularu, Gabriel Stanovsky:
ReliableEval: A Recipe for Stochastic LLM Evaluation via Method of Moments. CoRR abs/2505.22169 (2025) - [i5]Eliya Habba, Noam Dahan, Gili Lior, Gabriel Stanovsky:
PromptSuite: A Task-Agnostic Framework for Multi-Prompt Generation. CoRR abs/2507.14913 (2025) - [i4]Sarel Duanis, Asnat Greenstein-Messica, Eliya Habba:
JSON Whisperer: Efficient JSON Editing with LLMs. CoRR abs/2510.04717 (2025) - [i3]Anka Reuel, Avijit Ghosh, Jenny Chim, Andrew Tran, Yanan Long, Jennifer Mickel, Usman Gohar, Srishti Yadav, Pawan Sasanka Ammanamanchi, Mowafak Allaham, Hossein A. Rahmani, Mubashara Akhtar, Felix Friedrich, Robert Scholz, Michael Alexander Riegler, Jan Batzner, Eliya Habba, Arushi Saxena, Anastassia Kornilova, Kevin Wei, Prajna Soni, Yohan Mathew, Kevin Klyman, Jeba Sania, Subramanyam Sahoo, Olivia Beyer Bruvik, Pouya Sadeghi, Sujata S. Goswami, Angelina Wang, Yacine Jernite, Zeerak Talat, Stella Biderman, Mykel J. Kochenderfer, Sanmi Koyejo, Irene Solaiman:
Who Evaluates AI's Social Impacts? Mapping Coverage and Gaps in First and Third Party Evaluations. CoRR abs/2511.05613 (2025) - 2024
- [c2]Nitzan Bitton Guetta, Aviv Slobodkin, Aviya Maimon, Eliya Habba, Royi Rassin, Yonatan Bitton, Idan Szpektor, Amir Globerson, Yuval Elovici:
Visual Riddles: a Commonsense and World Knowledge Challenge for Large Vision and Language Models. NeurIPS 2024 - [i2]Nitzan Bitton Guetta, Aviv Slobodkin, Aviya Maimon, Eliya Habba, Royi Rassin, Yonatan Bitton, Idan Szpektor, Amir Globerson, Yuval Elovici:
Visual Riddles: a Commonsense and World Knowledge Challenge for Large Vision and Language Models. CoRR abs/2407.19474 (2024) - 2023
- [c1]Eliya Habba
, Renana Keydar
, Dan Bareket
, Gabriel Stanovsky
:
The Perfect Victim: Computational Analysis of Judicial Attitudes towards Victims of Sexual Violence. ICAIL 2023: 111-120 - [i1]Eliya Habba, Renana Keydar, Dan Bareket, Gabriel Stanovsky:
The Perfect Victim: Computational Analysis of Judicial Attitudes towards Victims of Sexual Violence. CoRR abs/2305.05302 (2023)
Coauthor Index
manage site settings
To protect your privacy, all features that rely on external API calls from your browser are turned off by default. You need to opt-in for them to become active. All settings here will be stored as cookies with your web browser. For more information see our F.A.Q.
Unpaywalled article links
Add open access links from to the list of external document links (if available).
Privacy notice: By enabling the option above, your browser will contact the API of unpaywall.org to load hyperlinks to open access articles. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Unpaywall privacy policy.
Archived links via Wayback Machine
For web page which are no longer available, try to retrieve content from the of the Internet Archive (if available).
Privacy notice: By enabling the option above, your browser will contact the API of archive.org to check for archived content of web pages that are no longer available. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Internet Archive privacy policy.
Reference lists
Add a list of references from ,
, and
to record detail pages.
load references from crossref.org and opencitations.net
Privacy notice: By enabling the option above, your browser will contact the APIs of crossref.org, opencitations.net, and semanticscholar.org to load article reference information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Crossref privacy policy and the OpenCitations privacy policy, as well as the AI2 Privacy Policy covering Semantic Scholar.
Citation data
Add a list of citing articles from and
to record detail pages.
load citations from opencitations.net
Privacy notice: By enabling the option above, your browser will contact the API of opencitations.net and semanticscholar.org to load citation information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the OpenCitations privacy policy as well as the AI2 Privacy Policy covering Semantic Scholar.
OpenAlex data
Load additional information about publications from .
Privacy notice: By enabling the option above, your browser will contact the API of openalex.org to load additional information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the information given by OpenAlex.
last updated on 2026-08-18 00:11 CEST by the dblp team
all metadata released as open data under CC0 1.0 license
see also: Terms of Use | Privacy Policy | Imprint