AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Model Releases

WeatherArchive-Bench: Benchmarking Retrieval-Augmented Reasoning for Historical Weather Archives

DGX agent

arXiv:2510.05336v2 Announce Type: replace Abstract: Historical archives on weather events are collections of enduring primary source records that offer rich, untapped narratives of how societies have

model-releasesarxiv-cs-cl
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

What Makes AI Research Replicable? Executable Knowledge Graphs as Scientific Knowledge Representations

DGX agent

arXiv:2510.17795v3 Announce Type: replace Abstract: Replicating AI research is a crucial yet challenging task for large language model (LLM) agents. Existing approaches often struggle to generate exec

agentsarxiv-cs-cl
21 Apr 2026
Research

What makes an entity salient in discourse?

DGX agent

arXiv:2508.16464v2 Announce Type: replace Abstract: Entities in discourse vary in salience: main participants, objects and locations stay prominent, while others are quickly forgotten, raising questio

researcharxiv-cs-cl
21 Apr 2026
Safety

When Choices Become Risks: Safety Failures of Large Language Models under Multiple-Choice Constraints

DGX agent

arXiv:2604.16916v1 Announce Type: new Abstract: Safety alignment in large language models (LLMs) is primarily evaluated under open-ended generation, where models can mitigate risk by refusing to respo

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

When Helpers Become Hazards: A Benchmark for Analyzing Multimodal LLM-Powered Safety in Daily Life

DGX agent

arXiv:2601.04043v2 Announce Type: replace Abstract: As Multimodal Large Language Models (MLLMs) become an indispensable assistant in human life, the unsafe content generated by MLLMs poses a danger to

model-releasesarxiv-cs-cl
21 Apr 2026
Research

When Informal Text Breaks NLI: Tokenization Failure, Distribution Shift, and Targeted Mitigations

DGX agent

arXiv:2604.16787v1 Announce Type: new Abstract: We study how informal surface forms degrade NLI accuracy in ELECTRA-small (14M) and RoBERTa-large (355M) across four transforms applied to SNLI and Mult

researcharxiv-cs-cl
21 Apr 2026
Research

When Misinformation Speaks and Converses: Rethinking Fact-Checking in Audio Platforms

DGX agent

arXiv:2604.16767v1 Announce Type: new Abstract: Audio platforms have evolved beyond entertainment. They have become central to public discourse, from podcasts and radio to WhatsApp voice notes and liv

researcharxiv-cs-cl
21 Apr 2026
Research

When More Words Say Less: Decoupling Length and Specificity in Image Description Evaluation

DGX agent

arXiv:2601.04609v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly used to make visual content accessible via text-based descriptions. In current systems, however, desc

researcharxiv-cs-cl
21 Apr 2026
Safety

Where Do Self-Supervised Speech Models Become Unfair?

DGX agent

arXiv:2604.18249v1 Announce Type: new Abstract: Speech encoder models are known to model members of some speaker groups (SGs) better than others. However, there has been little work in establishing wh

safetyarxiv-cs-cl
21 Apr 2026
Research

Where is the Mind? Persona Vectors and LLM Individuation

DGX agent

arXiv:2604.17031v1 Announce Type: new Abstract: The individuation problem for large language models asks which entities associated with them, if any, should be identified as minds. We approach this pr

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Who is the richest club in the championship? Detecting and Rewriting Underspecified Questions Improve QA Performance

DGX agent

arXiv:2602.11938v5 Announce Type: replace Abstract: Large language models (LLMs) perform well on well-posed questions, yet standard question-answering (QA) benchmarks remain far from solved. We argue

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Who Watches the Watchmen? Humans Disagree With Translation Metrics on Unseen Domains

DGX agent

arXiv:2604.17393v1 Announce Type: new Abstract: Automatic evaluation metrics are central to the development of machine translation systems, yet their robustness under domain shift remains unclear. Mos

researcharxiv-cs-cl
21 Apr 2026
Safety

Why Agents Compromise Safety Under Pressure

DGX agent

arXiv:2603.14975v2 Announce Type: replace-cross Abstract: Large Language Model agents deployed in complex environments frequently encounter a conflict between maximizing goal achievement and adhering

safetyarxiv-cs-cl
21 Apr 2026
Safety

Why AI Readiness Is an Organizational Learning Problem, Not a Technology Purchase

DGX agent

arXiv:2604.16369v1 Announce Type: cross Abstract: Global corporate AI investment reached $252.3 billion in 2024, yet only 6% of firms report significant earnings impact. This article argues that AI pr

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

WorldDB: A Vector Graph-of-Worlds Memory Engine with Ontology-Aware Write-Time Reconciliation

DGX agent

arXiv:2604.18478v1 Announce Type: cross Abstract: Persistent memory is the bottleneck separating stateless chatbots from long-running agentic systems. Retrieval-augmented generation (RAG) over flat ve

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Writing-RL: Advancing Long-form Writing via Adaptive Curriculum Reinforcement Learning

DGX agent

arXiv:2506.05760v2 Announce Type: replace Abstract: Recent advances in Large Language Models(LLMs) have enabled strong performance in long-form writing, but current training paradigms remain limited:

researcharxiv-cs-cl
21 Apr 2026
Research

x1: Learning to Think Adaptively Across Languages and Cultures

DGX agent

arXiv:2604.16917v1 Announce Type: new Abstract: Languages encode distinct abstractions and inductive priors, yet most large language models (LLMs) overlook this diversity by reasoning in a single domi

researcharxiv-cs-cl
21 Apr 2026
Safety

ZoFia: Zero-Shot Fake News Detection with Entity-Guided Retrieval and Multi-LLM Interaction

DGX agent

arXiv:2511.01188v2 Announce Type: replace Abstract: The rapid spread of fake news threatens social stability and public trust, highlighting the urgent need for its effective detection. Although large

safetyarxiv-cs-cl
21 Apr 2026
Tutorials

A Case Study on the Impact of Anonymization Along the RAG Pipeline

DGX agent

arXiv:2604.15958v1 Announce Type: cross Abstract: Despite the considerable promise of Retrieval-Augmented Generation (RAG), many real-world use cases may create privacy concerns, where the purported u

tutorialsarxiv-cs-cl
20 Apr 2026
Safety

A Systematic Study of Training-Free Methods for Trustworthy Large Language Models

DGX agent

arXiv:2604.15789v1 Announce Type: new Abstract: As Large Language Models (LLMs) receive increasing attention and are being deployed across various domains, their potential risks, including generating

safetyarxiv-cs-cl
20 Apr 2026
Research

Acoustic and Facial Markers of Perceived Conversational Success in Spontaneous Speech

DGX agent

arXiv:2604.15322v1 Announce Type: cross Abstract: Individuals often align their speaking patterns with their interlocutors, a phenomenon linked to engagement and rapport. While well documented in task

researcharxiv-cs-cl
20 Apr 2026
Model Releases

Aletheia: Gradient-Guided Layer Selection for Efficient LoRA Fine-Tuning Across Architectures

DGX agent

arXiv:2604.15351v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) has become the dominant parameter-efficient fine-tuning method for large language models, yet standard practice applies LoR

model-releasesarxiv-cs-cl
20 Apr 2026
Research

ATTNPO: Attention-Guided Process Supervision for Efficient Reasoning

DGX agent

arXiv:2602.09953v2 Announce Type: replace Abstract: Large reasoning models trained with reinforcement learning and verifiable rewards (RLVR) achieve strong performance on complex reasoning tasks, yet

researcharxiv-cs-cl
20 Apr 2026
Research

Author-in-the-Loop Response Generation and Evaluation: Integrating Author Expertise and Intent in Responses to Peer Review

DGX agent

arXiv:2602.11173v2 Announce Type: replace Abstract: Author response (rebuttal) writing is a critical stage of scientific peer review that demands substantial author effort. In practice, authors posses

researcharxiv-cs-cl
20 Apr 2026
Research

Brain Score Tracks Shared Properties of Languages: Evidence from Many Natural Languages and Structured Sequences

DGX agent

arXiv:2604.15503v1 Announce Type: new Abstract: Recent breakthroughs in language models (LMs) using neural networks have raised the question: how similar are these models' processing to human language

researcharxiv-cs-cl
20 Apr 2026
Safety

C-Mining: Unsupervised Discovery of Seeds for Cultural Data Synthesis via Geometric Misalignment

DGX agent

arXiv:2604.15675v1 Announce Type: new Abstract: Achieving cultural alignment in Large Language Models (LLMs) increasingly depends on synthetic data generation. For such synthesis, the most vital initi

safetyarxiv-cs-cl
20 Apr 2026
Model Releases

CHOP: Chunkwise Context-Preserving Framework for RAG on Multi Documents

DGX agent

arXiv:2604.15802v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems lose retrieval accuracy when similar documents coexist in the vector database, causing unnecessary informat

model-releasesarxiv-cs-cl
20 Apr 2026
Research

CIG: Measuring Conversational Information Gain in Deliberative Dialogues with Semantic Memory Dynamics

DGX agent

arXiv:2604.15647v1 Announce Type: new Abstract: Measuring the quality of public deliberation requires evaluating not only civility or argument structure, but also the informational progress of a conve

researcharxiv-cs-cl
20 Apr 2026
Research

CiPO: Counterfactual Unlearning for Large Reasoning Models through Iterative Preference Optimization

DGX agent

arXiv:2604.15847v1 Announce Type: new Abstract: Machine unlearning has gained increasing attention in recent years, as a promising technique to selectively remove unwanted privacy or copyrighted infor

researcharxiv-cs-cl
20 Apr 2026
Agents

CoEvolve: Training LLM Agents via Agent-Data Mutual Evolution

DGX agent

arXiv:2604.15840v1 Announce Type: new Abstract: Reinforcement learning for LLM agents is typically conducted on a static data distribution, which fails to adapt to the agent's evolving behavior and le

agentsarxiv-cs-cl
20 Apr 2026
Research

Collaboration of Fusion and Independence: Hypercomplex-driven Robust Multi-Modal Knowledge Graph Completion

DGX agent

arXiv:2509.23714v2 Announce Type: replace Abstract: Multi-modal knowledge graph completion (MMKGC) aims to discover missing facts in multi-modal knowledge graphs (MMKGs) by leveraging both structural

researcharxiv-cs-cl
20 Apr 2026
Model Releases

ConFu: Contemplate the Future for Better Speculative Sampling

DGX agent

arXiv:2603.08899v2 Announce Type: replace Abstract: Speculative decoding has emerged as a powerful approach to accelerate large language model (LLM) inference by employing lightweight draft models to

model-releasesarxiv-cs-cl
20 Apr 2026
Research

ConlangCrafter: Constructing Languages with a Multi-Hop LLM Pipeline

DGX agent

arXiv:2508.06094v4 Announce Type: replace Abstract: Constructed languages (conlangs) such as Esperanto and Quenya have played diverse roles in art, philosophy, and international communication. Meanwhi

researcharxiv-cs-cl
20 Apr 2026
Research

Creating and Evaluating Personas Using Generative AI: A Scoping Review of 81 Articles

DGX agent

arXiv:2504.04927v2 Announce Type: replace-cross Abstract: As generative AI (GenAI) is increasingly applied in persona development to represent real users, understanding the implications and limitation

researcharxiv-cs-cl
20 Apr 2026
Research

Curing Miracle Steps in LLM Mathematical Reasoning with Rubric Rewards

DGX agent

arXiv:2510.07774v3 Announce Type: replace Abstract: In this paper, we observe that current models are susceptible to reward hacking, leading to a substantial overestimation of a model's reasoning abil

researcharxiv-cs-cl
20 Apr 2026
Applications

Cut Your Losses! Learning to Prune Paths Early for Efficient Parallel Reasoning

DGX agent

arXiv:2604.16029v1 Announce Type: new Abstract: Parallel reasoning enhances Large Reasoning Models (LRMs) but incurs prohibitive costs due to futile paths caused by early errors. To mitigate this, pat

applicationsarxiv-cs-cl
20 Apr 2026
Applications

Designing Synthetic Discussion Generation Systems: A Case Study for Online Facilitation

DGX agent

arXiv:2503.16505v4 Announce Type: replace-cross Abstract: A critical challenge in social science research is the high cost associated with experiments involving human participants. We identify Synthet

applicationsarxiv-cs-cl
20 Apr 2026
Research

Detecting and Suppressing Reward Hacking with Gradient Fingerprints

DGX agent

arXiv:2604.16242v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) typically optimizes for outcome rewards without imposing constraints on intermediate reasoning.

researcharxiv-cs-cl
20 Apr 2026
Research

Disentangling Mathematical Reasoning in LLMs: A Methodological Investigation of Internal Mechanisms

DGX agent

arXiv:2604.15842v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated impressive capabilities, yet their internal mechanisms for handling reasoning-intensive tasks remain unde

researcharxiv-cs-cl
20 Apr 2026
Research

Do LLMs Really Know What They Don't Know? Internal States Mainly Reflect Knowledge Recall Rather Than Truthfulness

DGX agent

arXiv:2510.09033v3 Announce Type: replace Abstract: Recent work suggests that LLMs 'know what they don't know', positing that hallucinated and factually correct outputs arise from distinct internal pr

researcharxiv-cs-cl
20 Apr 2026
Model Releases

Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap

DGX agent

arXiv:2604.16256v1 Announce Type: cross Abstract: Reasoning in vision-language models (VLMs) has recently attracted significant attention due to its broad applicability across diverse downstream tasks

model-releasesarxiv-cs-cl
20 Apr 2026
Agents

Evaluating LLM Simulators as Differentially Private Data Generators

DGX agent

arXiv:2604.15461v1 Announce Type: cross Abstract: LLM-based simulators offer a promising path for generating complex synthetic data where traditional differentially private (DP) methods struggle with

agentsarxiv-cs-cl
20 Apr 2026
Model Releases

Exploring the Capability Boundaries of LLMs in Mastering of Chinese Chouxiang Language

DGX agent

arXiv:2604.15841v1 Announce Type: new Abstract: While large language models (LLMs) have achieved remarkable success in general language tasks, their performance on Chouxiang Language, a representative

model-releasesarxiv-cs-cl
20 Apr 2026
Agents

FACTS: Table Summarization via Offline Template Generation with Agentic Workflows

DGX agent

arXiv:2510.13920v2 Announce Type: replace Abstract: Query-focused table summarization requires generating natural language summaries of tabular data conditioned on a user query, enabling users to acce

agentsarxiv-cs-cl
20 Apr 2026
Research

Faithfulness-Aware Uncertainty Quantification for Fact-Checking the Output of Retrieval Augmented Generation

DGX agent

arXiv:2505.21072v4 Announce Type: replace Abstract: Large Language Models (LLMs) enhanced with retrieval, an approach known as Retrieval-Augmented Generation (RAG), have achieved strong performance in

researcharxiv-cs-cl
20 Apr 2026
Research

Faster LLM Inference via Sequential Monte Carlo

DGX agent

arXiv:2604.15672v1 Announce Type: cross Abstract: Speculative decoding (SD) accelerates language model inference by drafting tokens from a cheap proposal model and verifying them against an expensive

researcharxiv-cs-cl
20 Apr 2026
Research

FD-NL2SQL: Feedback-Driven Clinical NL2SQL that Improves with Use

DGX agent

arXiv:2604.15646v1 Announce Type: new Abstract: Clinicians exploring oncology trial repositories often need ad-hoc, multi-constraint queries over biomarkers, endpoints, interventions, and time, yet wr

researcharxiv-cs-cl
20 Apr 2026
Safety

Follow the Flow: On Information Flow Across Textual Tokens in Text-to-Image Models

DGX agent

arXiv:2504.01137v3 Announce Type: replace Abstract: Text-to-image generation models suffer from alignment problems, where generated images fail to accurately capture the objects and relations in the t

safetyarxiv-cs-cl
20 Apr 2026
← Previous
1…139140141142143…160
Next →