CompanyAnthropic8 recent entries22 May 2026Evaluating Commercial AI Chatbots as News IntermediariesarXiv:2605.22785v1 Announce Type: new Abstract: AI chatbots are rapidly shaping how people encounter the news, yet no prior study has systematically measured how accurately these systems, with their p→28 May 2026A Query Engine for the AgentsarXiv:2605.27785v1 Announce Type: new Abstract: The fastest-growing data in production today is unstructured text: agent traces, chat logs, reasoning chains, model outputs. People want to analyze it, →
CompanyOpenAI8 recent entries17 Apr 2026An Underexplored Frontier: Large Language Models for Rare Disease Patient Education and Communication -- A scoping reviewarXiv:2604.14179v1 Announce Type: new Abstract: Rare diseases affect over 300 million people worldwide and are characterized by complex care pathways, limited clinical expertise, and substantial unmet→21 Apr 2026Generative midtended cognition and Artificial Intelligence. Thinging with thinging thingsarXiv:2411.06812v2 Announce Type: replace-cross Abstract: This paper introduces the concept of ``generative midtended cognition'', exploring the integration of generative AI with human cognition. The →4 Jun 2026Stumbling Into AI Emotional Dependence: How Routine AI Interactions Reshape Human ConnectionarXiv:2606.04150v1 Announce Type: new Abstract: Public discourse and emerging policy typically assume that AI emotional support is a deliberate act: a lonely user consciously seeking comfort from a de→30 Jun 2026Insidious by Design: Implications of Large Language Model algorithmic bias for the Global SoutharXiv:2606.28333v1 Announce Type: cross Abstract: egin{quote} The biases in Large Language Models' (LLMs) outputs remain inadequately theorised, particularly from the perspective of the Global South. →8 Jul 2026Depression Symptoms and Relational Patterns in 187k ChatGPT HistoriesarXiv:2607.05685v1 Announce Type: cross Abstract: Large language models are increasingly used as private, always-available conversational systems, but little is known about how people with depressive →23 Jul 2026Are Attributions of Consciousness to AI Chatbots Epistemically Innocent?arXiv:2607.20001v1 Announce Type: cross Abstract: Artificial intelligence (AI) chatbots (e.g., ChatGPT) can communicate in strikingly humanlike ways. This has prompted many chatbot users to attribute →30 Jul 2026Identifying Implicit Bias in LLM-based Chat AI Toward People with Intellectual DisabilitiesarXiv:2607.26062v1 Announce Type: cross Abstract: Background: This work investigates the presence of implicit bias in Large Language Model (LLM)-based chat AI models directed toward people with intell→11 Aug 2026Weather- and Location-Aware Agentic Dining Recommendation: Leveraging LLM World Knowledge for Region-Sensitive Contextual ReasoningarXiv:2608.07593v1 Announce Type: cross Abstract: Context-aware recommender systems have long recognized that factors such as location, time, and weather shape where and what people choose to eat. Exi
CompanyGoogle8 recent entries22 May 2026Evaluating Commercial AI Chatbots as News IntermediariesarXiv:2605.22785v1 Announce Type: new Abstract: AI chatbots are rapidly shaping how people encounter the news, yet no prior study has systematically measured how accurately these systems, with their p→3 Jun 2026Generating the Modal Worker: A Cross-Model Audit of Race and Gender in LLM-Generated Personas Across 41 OccupationsarXiv:2510.21011v3 Announce Type: replace-cross Abstract: As generative AI tools are increasingly used to portray people in professional roles, understanding their racial and gender representational b→8 Jun 2026UrduMMLU: A Massive Multitask Benchmark for Urdu Language UnderstandingarXiv:2606.07167v1 Announce Type: cross Abstract: Meaningful multilingual evaluation must test models in the target language and educational context. Urdu, spoken by more than 230 million people, lack→11 Jun 2026The N-Body Problem: Parallel Execution from Single-Person Egocentric VideoarXiv:2512.11393v2 Announce Type: replace Abstract: Humans can intuitively parallelise complex activities, but can a model predict this from observing a single person? Given one egocentric video, we i→26 Jun 2026WatchAct: A Benchmark for Behavior-Grounded Robot ManipulationarXiv:2606.26443v1 Announce Type: cross Abstract: A robot working alongside people must reason about what they have done, in what order, and with what intent. Video carries the spatial layouts, object→3 Jul 2026VisionAId: An Offline-First Multimodal Android Assistant for People with Visual Impairment, Featuring Personalized Object RetrievalarXiv:2607.02371v1 Announce Type: cross Abstract: Over 285 million people worldwide live with a visual impairment, for whom everyday tasks such as avoiding obstacles, locating personal belongings, rec→28 Jul 2026Who Gets Named: Citation Type Predicts Individual Naming by Grounded Language Models, and a Roster Instrument Captures 0.5% of ItarXiv:2607.23893v1 Announce Type: cross Abstract: Prior work on AI brand visibility measures the firm: does a model recommend a company, and does that track its reputation. This study asks the questio→29 Jul 2026Personalization, Personas, and Forecasting in Value AlignmentarXiv:2607.24782v1 Announce Type: new Abstract: LLM behavior may be conditioned by human identity in several ways: they may be asked to adapt to users, role-play populations, or forecast how people wo
CompanyMeta8 recent entries12 May 2026K12-KGraph: A Curriculum-Aligned Knowledge Graph for Benchmarking and Training Educational LLMsarXiv:2605.09635v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in K-12 education, yet existing benchmarks such as C-Eval, CMMLU, GaokaoBench, and EduEval mainly eva→27 May 2026SIA: Self Improving AI with Harness & Weight UpdatesarXiv:2605.27276v1 Announce Type: new Abstract: Humans are the bottleneck in building and improving AI. Both the models and the agents that wrap them are written, tuned, and corrected by people. The l→5 Jun 2026From Self to Other: Evaluating Demographic Perspective-Taking in LLM Hate Speech AnnotationarXiv:2606.06266v1 Announce Type: new Abstract: Hate speech detection is inherently subjective: people from different demographic groups perceive the same content very differently. Collecting enough a→23 Jun 2026Measuring Intent Comprehension in LLMsarXiv:2506.16584v3 Announce Type: replace-cross Abstract: People judge interactions with large language models (LLMs) as successful when outputs match what they want, not what they type. Yet LLMs are →26 Jun 2026Assert, don't describe: Linguistic features that shift LLM reasoning about animal welfarearXiv:2606.26104v1 Announce Type: cross Abstract: Animal-welfare advocates produce a lot of writing, and increasingly that writing trains the language models that millions of people then ask about ani→7 Jul 2026IRC-Bench: Recognizing Entities from Contextual Cues in First-Person ReminiscencesarXiv:2605.06142v2 Announce Type: replace-cross Abstract: When people recount personal memories, they often refer to people, places, and events indirectly, relying on con-textual cues rather than expl→28 Jul 2026Simulating Tenant Responses to Energy Policy Interventions with Transaction-Cost-Aware LLM AgearXiv:2607.24341v1 Announce Type: new Abstract: Recent studies use Large language models (LLMs) to simulate human opinions and decisions by prompting models with demographic, attitudinal, or persona-b→30 Jul 2026Identifying Implicit Bias in LLM-based Chat AI Toward People with Intellectual DisabilitiesarXiv:2607.26062v1 Announce Type: cross Abstract: Background: This work investigates the presence of implicit bias in Large Language Model (LLM)-based chat AI models directed toward people with intell
CompanyMistral3 recent entries19 May 2026RTI-Bench: A Structured Dataset for Indian Right-to-Information Decision AnalysisarXiv:2605.16843v1 Announce Type: new Abstract: India's Right to Information Act, 2005 gives every citizen the right to demand information from public authorities, yet in practice most people cannot m→3 Jun 2026Generating the Modal Worker: A Cross-Model Audit of Race and Gender in LLM-Generated Personas Across 41 OccupationsarXiv:2510.21011v3 Announce Type: replace-cross Abstract: As generative AI tools are increasingly used to portray people in professional roles, understanding their racial and gender representational b→30 Jul 2026Identifying Implicit Bias in LLM-based Chat AI Toward People with Intellectual DisabilitiesarXiv:2607.26062v1 Announce Type: cross Abstract: Background: This work investigates the presence of implicit bias in Large Language Model (LLM)-based chat AI models directed toward people with intell
CompanyxAI4 recent entries22 May 2026Evaluating Commercial AI Chatbots as News IntermediariesarXiv:2605.22785v1 Announce Type: new Abstract: AI chatbots are rapidly shaping how people encounter the news, yet no prior study has systematically measured how accurately these systems, with their p→30 Jun 2026Insidious by Design: Implications of Large Language Model algorithmic bias for the Global SoutharXiv:2606.28333v1 Announce Type: cross Abstract: egin{quote} The biases in Large Language Models' (LLMs) outputs remain inadequately theorised, particularly from the perspective of the Global South. →15 Jul 2026The One-Word Census: Answer-Choice Conformity Across 44 Language ModelsarXiv:2607.12796v1 Announce Type: cross Abstract: When a language model must pick one answer from a large space of equally valid options, which does it pick -- and how often is it the same answer ever→28 Jul 2026Who Gets Named: Citation Type Predicts Individual Naming by Grounded Language Models, and a Roster Instrument Captures 0.5% of ItarXiv:2607.23893v1 Announce Type: cross Abstract: Prior work on AI brand visibility measures the firm: does a model recommend a company, and does that track its reputation. This study asks the questio
CompanyDeepSeek3 recent entries21 Apr 2026BengaliMoralBench: A Benchmark for Auditing Moral Reasoning in Large Language Models within Bengali Language and CulturearXiv:2511.03180v2 Announce Type: replace Abstract: As multilingual Large Language Models (LLMs) gain traction across South Asia, their alignment with local ethical norms, particularly for Bengali, sp→3 Jun 2026Generating the Modal Worker: A Cross-Model Audit of Race and Gender in LLM-Generated Personas Across 41 OccupationsarXiv:2510.21011v3 Announce Type: replace-cross Abstract: As generative AI tools are increasingly used to portray people in professional roles, understanding their racial and gender representational b→30 Jun 2026Cross-Temporal Sinhala OCR: Page-Level Adaptation and Diachronic AnalysisarXiv:2606.29378v1 Announce Type: new Abstract: Sinhala is a morphologically rich abugida spoken by roughly 16 million people in Sri Lanka, and to date, there are no publicly available real-world data
CompanyNVIDIA2 recent entries27 Apr 2026Kernel Contracts: A Specification Language for ML Kernel Correctness Across Heterogeneous SiliconarXiv:2604.22032v1 Announce Type: new Abstract: Every ML kernel ships with an implicit contract about what it computes. People rarely write the contract down. When two kernels disagree -- when a matmu→16 Jul 2026Text2Sign: A Single-GPU Diffusion Baseline for Text-to-Sign Language Video GenerationarXiv:2607.13164v1 Announce Type: cross Abstract: Sign language is a primary communication channel for millions of Deaf and hard-of-hearing people, yet text-to-signer video generation remains costly b