Publications

Sample-Size Scaling of the African Languages NLI Evaluation

Anuj Tiwari, Oluwapelumi Ogunremu, Terry Oko-odion, and 2 Others

AfricaNLP Workshop, EACL'26

From Script to Semantics: Prompting Strategies for African NLI

Anuj Tiwari, Terry Oko-odion, and Hannah Nwokocha

RAIL Workshop, LREC'26

Lingo_Research_Group at SemEval-2026 Task 9: Evaluating Prompt Variants for Polarization Detection

Pritam Kadasi†, Anuj Tiwari†, and Mayank Singh

SemEval Workshop, ACL'26

Chapter: Case Studies in Data-Driven Sports Medicine

Anuj Tiwari

Data-Driven Sports Medicine Book, Wiley Scrivener Publishing

Research Artifacts

Benchmarking Gemma-3 Series Models on EMMA-mini for Multimodal Reasoning

Anuj Tiwari

Code

Campus Food Habits & Delivery Preferences

Anuj Tiwari and Adeeba Irfan

Dataset

Hairstyle Survey

Anuj Tiwari

Dataset

Academic Services

  • Poster Reviewer: AI Across Cultures Workshop, CHI 2026
  • System Description Paper Reviewer: SemEval Workshop, ACL 2026
  • Research Paper Reviewer: RAIL Workshop, LREC 2026
  • Research Paper Reviewer: Main Track, IndabaXNigeria 2026

Recommendations

Pritam Kadasi
Pritam Kadasi Profile
AI Research Engineer at Logituit | PhD CS, IIT Gandhinagar
I managed Anuj directly
"Anuj designed the 12 prompt evaluation framework that is the basis of our SemEval Workshop work and also led the cross-lingual analysis explaining performance degradation across subtasks and languages. I'd highlight his instinct to look for reviewer concerns in particular, as he picked up on the majority of those concerns prior to submission, and the few that did come up were minor fixes."
Terry Oko-odion
Terry Oko-odion Profile
Prospective Graduate Student | BSc CS, AAU Ekpoma
I worked with Anuj on the same team
"I have worked with Anuj on various projects but here I want to highlight our EACL workshop paper in particular. Initially we were trying to expand AfriXNLI, but ran into problems getting the data from the native language speakers. Anuj took the lead by turning this constraint into a more interesting research question by asking whether more data even helps in low-resource evaluation, and if yes, how much? He did the experiments and found that the clearest evidence for a language specific vs model driven scaling behavior is observed between Yoruba and Kinyarwanda language. He set the direction for the project and then together we worked on it, which made the paper possible. That kind of pivot from a dead end to a more useful question is what makes Anuj worth collaborating with."