OpenAI Introduces IndQA Benchmark for Indian Languages and Culture
OpenAI has introduced IndQA, a new benchmark designed to measure how well AI systems understand Indian languages and culture, according to an announcement on the company’s website. The benchmark aims to evaluate reasoning and contextual understanding across a broad range of cultural topics rather than focusing solely on translation or multiple-choice tasks.
IndQA includes 2,278 expert-authored questions in 12 Indian languages, covering 10 cultural domains such as literature, food, history, law, and religion. The dataset was created with contributions from 261 Indian experts, including linguists, journalists, artists, and scholars. Each question includes a rubric for grading responses, an ideal answer, and an English translation for consistency.
Questions were tested against several of OpenAI’s advanced models, including GPT‑4o, OpenAI o3, GPT‑4.5, and GPT‑5, to ensure difficulty and room for improvement. Evaluation results show that the GPT‑5 Thinking High model achieved the highest IndQA score among current models, though performance varied across languages and domains.
OpenAI stated that IndQA will be used to track progress in multilingual and culturally grounded reasoning and may serve as a model for similar benchmarks in other regions and languages.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like Daily AI Brief.
Also, consider following us on social media:
Daily AI Brief
Daily report covering major AI developments and industry news, with both top stories and complete market updates
Market report
2025 Generative AI in Professional Services Report
Thomson Reuters
This report by Thomson Reuters explores the integration and impact of generative AI technologies, such as ChatGPT and Microsoft Copilot, within the professional services sector. It highlights the growing adoption of GenAI tools across industries like legal, tax, accounting, and government, and discusses the challenges and opportunities these technologies present. The report also examines professionals' perceptions of GenAI and the need for strategic integration to maximize its value.
Read moreYou may also like
Appier Studies How AI Detects Missing Answers and Selects Reasoning Languages
EAIGG and Draup Introduce National AI Readiness Scorecard
Artmarket.com Tests How Five AI Systems Interpret the Same Book
Omneky Introduces TASTE BENCH for AI Ad Evaluation
Overmind Open Sources Platform for Specialized AI Models
Daily AI Brief: the AI news that matters, in your inbox.