paper-with-me

Papers

Khana: A Comprehensive Indian Cuisine Dataset

2025-09-07 · Omkar Prabhu arxiv

As global interest in diverse culinary experiences grows, food image models are essential for improving food-related applications by enabling accurate food recognition, recipe suggestions, dietary tracking, and automated meal planning. Despite the abundance of food datasets, a noticeable gap remains in capturing the nuances of Indian cuisine due to its vast regional diversity, complex preparations, and the lack of comprehensive labeled datasets that cover its full breadth. Through this exploration, we uncover Khana, a new benchmark dataset for food image classification, segmentation, and retrieval of dishes from Indian cuisine. Khana fills the gap by establishing a taxonomy of Indian cuisine and offering around 131K images in the dataset spread across 80 labels, each with a resolution of 500x500 pixels. This paper describes the dataset creation process and evaluates state-of-the-art models on classification, segmentation, and retrieval as baselines. Khana bridges the gap between research and development by providing a comprehensive and challenging benchmark for researchers while also serving as a valuable resource for developers creating real-world applications that leverage the rich tapestry of Indian cuisine. Webpage: https://khana.omkar.xyz

📄 PDF Abstract BibTeX arXiv:2509.06006

Code (0)

등록된 구현이 없습니다.

Tasks

Image Classification

Similar Papers 제목 키워드 기반

Othering and low status framing of immigrant cuisines in US restaurant reviews and large language models

2023-07-14 · Yiwei Luo, Kristina Gligorić, Dan Jurafsky

Identifying implicit attitudes toward food can mitigate social prejudice due to food's salience as a marker of ethnic identity. Stereotypes about food are representational harms that may contribute to racialized discours…

Text Generation

An Investigation of Linguistic Biases in LLM-Based Recommendations

2026-04-28 · Nitin Venkateswaran, Jason Ang, Deep Adhikari, Tarun Krishna Dasari arxiv

We investigate linguistic biases in LLM-based restaurant and product recommendations given prompts varying across Southern American English (AE), Indian English (IE), and Code-Switched Hindi-English dialects, using the Y…

Thought-For-Food: Reasoning Chain Induced Food Visual Question Answering

2025-11-03 · Riddhi Jain, Manasi Patwardhan, Parijat Deshpande, Venkataramana Runkana arxiv

The immense diversity in the culture and culinary of Indian cuisines calls attention to the major shortcoming of the existing Visual Question Answering(VQA) systems which are inclined towards the foods from Western regio…

Visual Question AnsweringReinforcement Learning

Automatic Speech Recognition Advancements for Indigenous Languages of the Americas

2024-04-12 · Monica Romero, Sandra Gomez, Ivan G. Torre

Indigenous languages are a fundamental legacy in the development of human communication, embodying the unique identity and culture of local communities in America. The Second AmericasNLP (Americas Natural Language Proces…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentationspeech-recognition+1

DRISHTIKON: A Multimodal Multilingual Benchmark for Testing Language Models' Understanding on Indian Culture

2025-09-23 · Arijit Maji, Raghvendra Kumar, Akash Ghosh, Anushka 외 arxiv

We introduce DRISHTIKON, a first-of-its-kind multimodal and multilingual benchmark centered exclusively on Indian culture, designed to evaluate the cultural understanding of generative AI systems. Unlike existing benchma…