AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Safety

When Bigger Isn't Better: A Comprehensive Fairness Evaluation of Political Bias in Multi-News Summarisation

DGX agent

arXiv:2604.21309v1 Announce Type: new Abstract: Multi-document news summarisation systems are increasingly adopted for their convenience in processing vast daily news content, making fairness across d

safetyarxiv-cs-cl
24 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

XtraGPT: Context-Aware and Controllable Academic Paper Revision via Human-AI Collaboration

DGX agent

arXiv:2505.11336v4 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in academic workflows, their capabilities remain limited in supporting high-quality sci

safetyarxiv-cs-cl
24 Apr 2026
Research

AFMRL: Attribute-Enhanced Fine-Grained Multi-Modal Representation Learning in E-commerce

DGX agent

arXiv:2604.20135v1 Announce Type: new Abstract: Multimodal representation is crucial for E-commerce tasks such as identical product retrieval. Large representation models (e.g., VLM2Vec) demonstrate s

researcharxiv-cs-cl
23 Apr 2026
Safety

Aligning Human-AI-Interaction Trust for Mental Health Support: Survey and Position for Multi-Stakeholders

DGX agent

arXiv:2604.20166v1 Announce Type: new Abstract: Building trustworthy AI systems for mental health support is a shared priority across stakeholders from multiple disciplines. However, 'trustworthy' rem

safetyarxiv-cs-cl
23 Apr 2026
Research

Aligning Stuttered-Speech Research with End-User Needs: Scoping Review, Survey, and Guidelines

DGX agent

arXiv:2604.20535v1 Announce Type: new Abstract: Atypical speech is receiving greater attention in speech technology research, but much of this work unfolds with limited interdisciplinary dialogue. For

researcharxiv-cs-cl
23 Apr 2026
Safety

All Languages Matter: Understanding and Mitigating Language Bias in Multilingual RAG

DGX agent

arXiv:2604.20199v1 Announce Type: new Abstract: Multilingual Retrieval-Augmented Generation (mRAG) leverages cross-lingual evidence to ground Large Language Models (LLMs) in global knowledge. However,

safetyarxiv-cs-cl
23 Apr 2026
Model Releases

Are LLM Uncertainty and Correctness Encoded by the Same Features? A Functional Dissociation via Sparse Autoencoders

DGX agent

arXiv:2604.19974v1 Announce Type: cross Abstract: Large language models can be uncertain yet correct, or confident yet wrong, raising the question of whether their output-level uncertainty and their a

model-releasesarxiv-cs-cl
23 Apr 2026
Safety

Ask Only When Needed: Proactive Retrieval from Memory and Skills for Experience-Driven Lifelong Agents

DGX agent

arXiv:2604.20572v1 Announce Type: new Abstract: Online lifelong learning enables agents to accumulate experience across interactions and continually improve on long-horizon tasks. However, existing me

safetyarxiv-cs-cl
23 Apr 2026
Safety

Avoiding Overthinking and Underthinking: Curriculum-Aware Budget Scheduling for LLMs

DGX agent

arXiv:2604.19780v1 Announce Type: new Abstract: Scaling test-time compute via extended reasoning has become a key paradigm for improving the capabilities of large language models (LLMs). However, exis

safetyarxiv-cs-cl
23 Apr 2026
Research

Believing without Seeing: Quality Scores for Contextualizing Vision-Language Model Explanations

DGX agent

arXiv:2509.25844v3 Announce Type: replace Abstract: When people query Vision-Language Models (VLMs) but cannot see the accompanying visual context (e.g. for blind and low-vision users), augmenting VLM

researcharxiv-cs-cl
23 Apr 2026
Model Releases

Beyond Majority Voting: Towards Fine-grained and More Reliable Reward Signal for Test-Time Reinforcement Learning

DGX agent

arXiv:2512.15146v3 Announce Type: replace Abstract: Test-time reinforcement learning mitigates the reliance on annotated data by using majority voting results as pseudo-labels, emerging as a complemen

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Beyond the Crowd: LLM-Augmented Community Notes for Governing Health Misinformation

DGX agent

arXiv:2510.11423v3 Announce Type: replace-cross Abstract: Community Notes, the crowd-sourced misinformation governance system on X (formerly Twitter), allows users to flag misleading posts, attach con

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Bootstrapping Post-training Signals for Open-ended Tasks via Rubric-based Self-play on Pre-training Text

DGX agent

arXiv:2604.20051v1 Announce Type: new Abstract: Self-play has recently emerged as a promising paradigm to train Large Language Models (LLMs). In self-play, the target LLM creates the task input (e.g.,

model-releasesarxiv-cs-cl
23 Apr 2026
Safety

Breaking the Assistant Mold: Modeling Behavioral Variation in LLM Based Procedural Character Generation

DGX agent

arXiv:2601.03396v2 Announce Type: replace Abstract: Procedural content generation has enabled vast virtual worlds through levels, maps, and quests, but large-scale character generation remains underex

safetyarxiv-cs-cl
23 Apr 2026
Model Releases

Chasing the Public Score: User Pressure and Evaluation Exploitation in Coding Agent Workflows

DGX agent

arXiv:2604.20200v1 Announce Type: new Abstract: Frontier coding agents are increasingly used in workflows where users supervise progress primarily through repeated improvement of a public score, namel

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

CLIP-SVD: Efficient and Interpretable Vision-Language Adaptation via Singular Values

DGX agent

arXiv:2509.03740v3 Announce Type: replace-cross Abstract: Vision-language models (VLMs) like CLIP have shown impressive zero-shot and few-shot learning capabilities across diverse applications. Howeve

model-releasesarxiv-cs-cl
23 Apr 2026
Research

Commonsense Knowledge with Negation: A Resource to Enhance Negation Understanding

DGX agent

arXiv:2604.19921v1 Announce Type: new Abstract: Negation is a common and important semantic feature in natural language, yet Large Language Models (LLMs) struggle when negation is involved in natural

researcharxiv-cs-cl
23 Apr 2026
Research

Composition-RL: Compose Your Verifiable Prompts for Reinforcement Learning of Large Language Models

DGX agent

arXiv:2602.12036v2 Announce Type: replace Abstract: Large-scale verifiable prompts underpin the success of Reinforcement Learning with Verifiable Rewards (RLVR), but they contain many uninformative ex

researcharxiv-cs-cl
23 Apr 2026
Research

Construction of a Battery Research Knowledge Graph using a Global Open Catalog

DGX agent

arXiv:2604.20241v1 Announce Type: new Abstract: Battery research is a rapidly growing and highly interdisciplinary field, making it increasingly difficult to track relevant expertise and identify pote

researcharxiv-cs-cl
23 Apr 2026
Applications

Continuous Semantic Caching for Low-Cost LLM Serving

DGX agent

arXiv:2604.20021v1 Announce Type: cross Abstract: As Large Language Models (LLMs) become increasingly popular, caching responses so that they can be reused by users with semantically similar queries h

applicationsarxiv-cs-cl
23 Apr 2026
Model Releases

Cooperative Profiles Predict Multi-Agent LLM Team Performance in AI for Science Workflows

DGX agent

arXiv:2604.20658v1 Announce Type: new Abstract: Multi-agent systems built from teams of large language models (LLMs) are increasingly deployed for collaborative scientific reasoning and problem-solvin

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

CRAFT: Training-Free Cascaded Retrieval for Tabular QA

DGX agent

arXiv:2505.14984v2 Announce Type: replace Abstract: Open-Domain Table Question Answering (TQA) involves retrieving relevant tables from a large corpus to answer natural language queries. Traditional d

model-releasesarxiv-cs-cl
23 Apr 2026
Local Ai

Decoding Text Spans for Efficient and Accurate Named-Entity Recognition

DGX agent

arXiv:2604.20447v1 Announce Type: new Abstract: Named Entity Recognition (NER) is a key component in industrial information extraction pipelines, where systems must satisfy strict latency and throughp

local-aiarxiv-cs-cl
23 Apr 2026
Model Releases

Development and Preliminary Evaluation of a Domain-Specific Large Language Model for Tuberculosis Care in South Africa

DGX agent

arXiv:2604.19776v1 Announce Type: new Abstract: Tuberculosis (TB) is one of the world's deadliest infectious diseases, and in South Africa, it contributes a significant burden to the country's health

model-releasesarxiv-cs-cl
23 Apr 2026
Safety

DRIV-EX: Counterfactual Explanations for Driving LLMs

DGX agent

arXiv:2603.00696v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used as reasoning engines in autonomous driving, yet their decision-making remains opaque. We propose

safetyarxiv-cs-cl
23 Apr 2026
Agents

Dual-Cluster Memory Agent: Resolving Multi-Paradigm Ambiguity in Optimization Problem Solving

DGX agent

arXiv:2604.20183v1 Announce Type: new Abstract: Large Language Models (LLMs) often struggle with structural ambiguity in optimization problems, where a single problem admits multiple related but confl

agentsarxiv-cs-cl
23 Apr 2026
Model Releases

Duluth at SemEval-2026 Task 6: DeBERTa with LLM-Augmented Data for Unmasking Political Question Evasions

DGX agent

arXiv:2604.20168v1 Announce Type: new Abstract: This paper presents the Duluth approach to SemEval-2026 Task 6 on CLARITY: Unmasking Political Question Evasions. We address Task 1 (clarity-level class

model-releasesarxiv-cs-cl
23 Apr 2026
Research

Effects of Cross-lingual Evidence in Multilingual Medical Question Answering

DGX agent

arXiv:2604.20531v1 Announce Type: new Abstract: This paper investigates Multilingual Medical Question Answering across high-resource (English, Spanish, French, Italian) and low-resource (Basque, Kazak

researcharxiv-cs-cl
23 Apr 2026
Research

ESGLens: An LLM-Based RAG Framework for Interactive ESG Report Analysis and Score Prediction

DGX agent

arXiv:2604.19779v1 Announce Type: new Abstract: Environmental, Social, and Governance (ESG) reports are central to investment decision-making, yet their length, heterogeneous content, and lack of stan

researcharxiv-cs-cl
23 Apr 2026
Model Releases

Evidence of Layered Positional and Directional Constraints in the Voynich Manuscript: Implications for Cipher-Like Structure

DGX agent

arXiv:2604.19762v1 Announce Type: new Abstract: The Voynich Manuscript (VMS) exhibits a script of uncertain origin whose grapheme sequences have resisted linguistic analysis. We present a systematic a

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Finding Duplicates in 1.1M BDD Steps: cukereuse, a Paraphrase-Robust Static Detector for Cucumber and Gherkin

DGX agent

arXiv:2604.20462v1 Announce Type: cross Abstract: Behaviour-Driven Development (BDD) suites accumulate step-text duplication whose maintenance cost is established in prior work. Existing detection tec

model-releasesarxiv-cs-cl
23 Apr 2026
Tutorials

Foundational Design Principles and Patterns for Building Robust and Adaptive GenAI-Native Systems

DGX agent

arXiv:2508.15411v3 Announce Type: replace-cross Abstract: Generative AI (GenAI) has emerged as a transformative technology, demonstrating remarkable capabilities across diverse application domains. Ho

tutorialsarxiv-cs-cl
23 Apr 2026
Model Releases

From Recall to Forgetting: Benchmarking Long-Term Memory for Personalized Agents

DGX agent

arXiv:2604.20006v1 Announce Type: new Abstract: Personalized agents that interact with users over long periods must maintain persistent memory across sessions and update it as circumstances change. Ho

model-releasesarxiv-cs-cl
23 Apr 2026
Safety

Graph2Counsel: Clinically Grounded Synthetic Counseling Dialogue Generation from Client Psychological Graphs

DGX agent

arXiv:2604.20382v1 Announce Type: new Abstract: Rising demand for mental health support has increased interest in using Large Language Models (LLMs) for counseling. However, adapting LLMs to this high

safetyarxiv-cs-cl
23 Apr 2026
Agents

HaS: Accelerating RAG through Homology-Aware Speculative Retrieval

DGX agent

arXiv:2604.20452v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) expands the knowledge boundary of large language models (LLMs) at inference by retrieving external documents as c

agentsarxiv-cs-cl
23 Apr 2026
Model Releases

How Much Does Persuasion Strategy Matter? LLM-Annotated Evidence from Charitable Donation Dialogues

DGX agent

arXiv:2604.19783v1 Announce Type: new Abstract: Which persuasion strategies, if any, are associated with donation compliance? Answering this requires fine-grained strategy labels across a full corpus

model-releasesarxiv-cs-cl
23 Apr 2026
Tutorials

How to measure the optimality of word or gesture order with respect to the principle of swap distance minimization

DGX agent

arXiv:2604.01938v3 Announce Type: replace Abstract: The structure of all the permutations of a sequence can be represented as a permutohedron, a graph where vertices are permutations and two vertices

tutorialsarxiv-cs-cl
23 Apr 2026
Research

HumorRank: A Tournament-Based Leaderboard for Evaluating Humor Generation in Large Language Models

DGX agent

arXiv:2604.19786v1 Announce Type: new Abstract: Evaluating humor in large language models (LLMs) is an open challenge because existing approaches yield isolated, incomparable metrics rather than unifi

researcharxiv-cs-cl
23 Apr 2026
Model Releases

Hybrid Multi-Phase Page Matching and Multi-Layer Diff Detection for Japanese Building Permit Document Review

DGX agent

arXiv:2604.19770v1 Announce Type: new Abstract: We present a hybrid multi-phase page matching algorithm for automated comparison of Japanese building permit document sets. Building permit review in Ja

model-releasesarxiv-cs-cl
23 Apr 2026
Research

Improving End-to-End Training of Retrieval-Augmented Generation Models via Joint Stochastic Approximation

DGX agent

arXiv:2508.18168v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) has become a widely recognized paradigm to combine parametric memory with non-parametric memories. An RAG model

researcharxiv-cs-cl
23 Apr 2026
Applications

Interpretability from the Ground Up: Stakeholder-Centric Design of Automated Scoring in Educational Assessments

DGX agent

arXiv:2511.17069v3 Announce Type: replace Abstract: AI-driven automated scoring systems offer scalable and efficient means of evaluating complex student-generated responses. Yet, despite increasing de

applicationsarxiv-cs-cl
23 Apr 2026
Model Releases

Intersectional Fairness in Large Language Models

DGX agent

arXiv:2604.20677v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in socially sensitive settings, raising concerns about fairness and biases, particularly across i

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Knapsack Optimization-based Schema Linking for LLM-based Text-to-SQL Generation

DGX agent

arXiv:2502.12911v3 Announce Type: replace Abstract: Generating SQLs from user queries is a long-standing challenge, where the accuracy of initial schema linking significantly impacts subsequent SQL ge

model-releasesarxiv-cs-cl
23 Apr 2026
Safety

Language-Coupled Reinforcement Learning for Multilingual Retrieval-Augmented Generation

DGX agent

arXiv:2601.14896v2 Announce Type: replace Abstract: Multilingual retrieval-augmented generation (MRAG) requires models to effectively acquire and integrate beneficial external knowledge from multiling

safetyarxiv-cs-cl
23 Apr 2026
Safety

Large language models perceive cities through a culturally uneven baseline

DGX agent

arXiv:2604.20048v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to describe, evaluate and interpret places, yet it remains unclear whether they do so from a cultural

safetyarxiv-cs-cl
23 Apr 2026
Model Releases

Less Languages, Less Tokens: An Efficient Unified Logic Cross-lingual Chain-of-Thought Reasoning Framework

DGX agent

arXiv:2604.20090v1 Announce Type: new Abstract: Cross-lingual chain-of-thought (XCoT) with self-consistency markedly enhances multilingual reasoning, yet existing methods remain costly due to extensiv

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

LLAMADRS: Evaluating Open-Source LLMs on Real Clinical Interviews--To Reason or Not to Reason?

DGX agent

arXiv:2501.03624v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel on many NLP benchmarks, but their behavior on real-world, semi-structured prediction remains underexplored.

model-releasesarxiv-cs-cl
23 Apr 2026
Research

LLM StructCore: Schema-Guided Reasoning Condensation and Deterministic Compilation

DGX agent

arXiv:2604.20560v1 Announce Type: new Abstract: Automatically filling Case Report Forms (CRFs) from clinical notes is challenging due to noisy language, strict output contracts, and the high cost of f

researcharxiv-cs-cl
23 Apr 2026
← Previous
1…126127128129130…161
Next →