AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Research

A Computational Method for Measuring 'Open Codes' in Qualitative Analysis

DGX agent

arXiv:2411.12142v4 Announce Type: replace Abstract: Qualitative analysis is critical to understanding human datasets in many social science disciplines. A central method in this process is inductive c

researcharxiv-cs-cl
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

A Goal Without a Plan Is Just a Wish: Efficient and Effective Global Planner Training for Long-Horizon Agent Tasks

DGX agent

arXiv:2510.05608v2 Announce Type: replace Abstract: Agents based on large language models (LLMs) struggle with brainless trial-and-error and generating hallucinatory actions due to a lack of global pl

agentsarxiv-cs-cl
21 Apr 2026
Agents

A Multi-Agent Approach for Claim Verification from Tabular Data Documents

DGX agent

arXiv:2604.17225v1 Announce Type: new Abstract: We present a novel approach for claim verification from tabular data documents. Recent LLM-based approaches either employ complex pretraining/fine-tunin

agentsarxiv-cs-cl
21 Apr 2026
Applications

A multimodal and temporal foundation model for virtual patient representations at healthcare system scale

DGX agent

arXiv:2604.18570v1 Announce Type: cross Abstract: Modern medicine generates vast multimodal data across siloed systems, yet no existing model integrates the full breadth and temporal depth of the clin

applicationsarxiv-cs-cl
21 Apr 2026
Research

A novel LSTM music generator based on the fractional time-frequency feature extraction

DGX agent

arXiv:2604.17823v1 Announce Type: cross Abstract: In this paper, we propose a novel approach for generating music based on an artificial intelligence (AI) system. We analyze the features of music and

researcharxiv-cs-cl
21 Apr 2026
Agents

A Survey on the Security of Long-Term Memory in LLM Agents: Toward Mnemonic Sovereignty

DGX agent

arXiv:2604.16548v1 Announce Type: cross Abstract: Research on large language model (LLM) security is shifting from 'will the model leak training data' to a more consequential question: can an agent wi

agentsarxiv-cs-cl
21 Apr 2026
Safety

A Text-To-Text Alignment Algorithm for Better Evaluation of Modern Speech Recognition Systems

DGX agent

arXiv:2509.24478v2 Announce Type: replace Abstract: Modern neural networks have greatly improved performance across speech recognition benchmarks. However, gains are often driven by frequent words wit

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

A Transformer and Prototype-based Interpretable Model for Contextual Sarcasm Detection

DGX agent

arXiv:2503.11838v2 Announce Type: replace Abstract: Sarcasm detection, with its figurative nature, poses unique challenges for affective systems designed to perform sentiment analysis. While these sys

model-releasesarxiv-cs-cl
21 Apr 2026
Research

A Universal Avoidance Method for Diverse Multi-branch Generation

DGX agent

arXiv:2604.17323v1 Announce Type: new Abstract: Modern generative models still lack human-level creativity, particularly in multi-branch diversity. Prior approaches to address this problem often incur

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Abstain-R1: Calibrated Abstention and Post-Refusal Clarification via Verifiable RL

DGX agent

arXiv:2604.17073v1 Announce Type: new Abstract: Reinforcement fine-tuning improves the reasoning ability of large language models, but it can also encourage them to answer unanswerable queries by gues

model-releasesarxiv-cs-cl
21 Apr 2026
Agents

AdaExplore: Failure-Driven Adaptation and Diversity-Preserving Search for Efficient Kernel Generation

DGX agent

arXiv:2604.16625v1 Announce Type: new Abstract: Recent large language model (LLM) agents have shown promise in using execution feedback for test-time adaptation. However, robust self-improvement remai

agentsarxiv-cs-cl
21 Apr 2026
Model Releases

Adaptive Text Anonymization: Learning Privacy-Utility Trade-offs via Prompt Optimization

DGX agent

arXiv:2602.20743v2 Announce Type: replace Abstract: Anonymizing textual documents is a highly context-sensitive problem: the appropriate balance between privacy protection and utility preservation var

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Adversarial Humanities Benchmark: Results on Stylistic Robustness in Frontier Model Safety

DGX agent

arXiv:2604.18487v1 Announce Type: new Abstract: The Adversarial Humanities Benchmark (AHB) evaluates whether model safety refusals survive a shift away from familiar harmful prompt forms. Starting fro

model-releasesarxiv-cs-cl
21 Apr 2026
Agents

Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence

DGX agent

arXiv:2604.18292v1 Announce Type: cross Abstract: Large language models are increasingly expected to serve as general-purpose agents that interact with external, stateful tool environments. The Model

agentsarxiv-cs-cl
21 Apr 2026
Agents

Agents Explore but Agents Ignore: LLMs Lack Environmental Curiosity

DGX agent

arXiv:2604.17609v1 Announce Type: new Abstract: LLM-based agents are assumed to integrate environmental observations into their reasoning: discovering highly relevant but unexpected information should

agentsarxiv-cs-cl
21 Apr 2026
Safety

Agree, Disagree, Explain: Decomposing Human Label Variation in NLI through the Lens of Explanations

DGX agent

arXiv:2510.16458v2 Announce Type: replace Abstract: Natural Language Inference (NLI) datasets often exhibit human label variation. To better understand these variations, explanation-based approaches a

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Alexandria: A Multi-Domain Dialectal Arabic Machine Translation Dataset for Culturally Inclusive and Linguistically Diverse LLMs

DGX agent

arXiv:2601.13099v2 Announce Type: replace Abstract: Arabic is a highly diglossic language where most daily communication occurs in regional dialects rather than Modern Standard Arabic (MSA). Despite t

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Align Documents to Questions: Question-Oriented Document Rewriting for Retrieval-Augmented Generation

DGX agent

arXiv:2604.17325v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) enhances the factuality of Large Language Models (LLMs) by incorporating retrieved documents and/or generated conte

safetyarxiv-cs-cl
21 Apr 2026
Safety

Aligning Backchannel and Dialogue Context Representations via Contrastive LLM Fine-Tuning

DGX agent

arXiv:2604.16622v1 Announce Type: new Abstract: Backchannels (e.g., `yeah', `mhm', and `right') are short, non-interruptive feedback signals whose lexical form and prosody jointly convey pragmatic mea

safetyarxiv-cs-cl
21 Apr 2026
Safety

Aligning Language Models for Lyric-to-Melody Generation with Rule-Based Musical Constraints

DGX agent

arXiv:2604.18489v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise in lyric-to-melody generation, but models trained with Supervised Fine-Tuning (SFT) often produce musically

safetyarxiv-cs-cl
21 Apr 2026
Local Ai

Aligning Language Models with Real-time Knowledge Editing

DGX agent

arXiv:2508.01302v3 Announce Type: replace Abstract: Knowledge editing aims to modify outdated knowledge in language models efficiently while retaining their original capabilities. Mainstream datasets

local-aiarxiv-cs-cl
21 Apr 2026
Safety

Alignment Data Map for Efficient Preference Data Selection and Diagnosis

DGX agent

arXiv:2505.23114v3 Announce Type: replace Abstract: Human preference data is essential for aligning large language models (LLMs) with human values, but collecting such data is often costly and ineffic

safetyarxiv-cs-cl
21 Apr 2026
Applications

AlphaContext: An Evolutionary Tree-based Psychometric Context Generator for Creativity Assessment

DGX agent

arXiv:2604.18398v1 Announce Type: new Abstract: Creativity has become a core competence in the era of LLMs and human-AI collaboration, underpinning innovation in real-world problem solving. Crucially,

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

Althea: Human-AI Collaboration for Fact-Checking and Critical Reasoning

DGX agent

arXiv:2602.11161v2 Announce Type: replace-cross Abstract: The web's information ecosystem demands fact-checking systems that are both scalable and epistemically trustworthy. Automated approaches offer

model-releasesarxiv-cs-cl
21 Apr 2026
Research

An Existence Proof for Neural Language Models That Can Explain Garden-Path Effects via Surprisal

DGX agent

arXiv:2604.18293v1 Announce Type: new Abstract: Surprisal theory hypothesizes that the difficulty of human sentence processing increases linearly with surprisal, the negative log-probability of a word

researcharxiv-cs-cl
21 Apr 2026
Research

An Exploration of Mamba for Speech Self-Supervised Models

DGX agent

arXiv:2506.12606v2 Announce Type: replace Abstract: While Mamba has demonstrated strong performance in language modeling, its potential as a speech self-supervised learning (SSL) model remains underex

researcharxiv-cs-cl
21 Apr 2026
Model Releases

AnchorMem: Anchored Facts with Associative Contexts for Building Memory in Large Language Models

DGX agent

arXiv:2604.17377v1 Announce Type: new Abstract: While large language models have achieved remarkable performance in complex tasks, they still need a memory system to utilize historical experience in l

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Annotating Dimensions of Social Perception in Text: A Sentence-Level Dataset of Warmth and Competence

DGX agent

arXiv:2601.06316v2 Announce Type: replace Abstract: Warmth (W) (often further broken down intoTrust (T) and Sociability (S)) and Competence (C) are central dimensions along which people evaluate indiv

researcharxiv-cs-cl
21 Apr 2026
Research

Annotation Entropy Predicts Per-Example Learning Dynamics in LoRA Fine-Tuning

DGX agent

arXiv:2604.16332v1 Announce Type: cross Abstract: We find that LoRA fine-tuning exhibits un-learning on contested examples: items with high annotator disagreement show increasing loss during training,

researcharxiv-cs-cl
21 Apr 2026
Agents

Answer Only as Precisely as Justified: Calibrated Claim-Level Specificity Control for Agentic Systems

DGX agent

arXiv:2604.17487v1 Announce Type: new Abstract: Agentic systems often fail not by being entirely wrong, but by being too precise: a response may be generally useful while particular claims exceed what

agentsarxiv-cs-cl
21 Apr 2026
Research

ArbGraph: Conflict-Aware Evidence Arbitration for Reliable Long-Form Retrieval-Augmented Generation

DGX agent

arXiv:2604.18362v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) remains unreliable in long-form settings, where retrieved evidence is noisy or contradictory, making it difficult f

researcharxiv-cs-cl
21 Apr 2026
Safety

Arch: An AI-Native Hardware Description Language for Register-Transfer Clocked Hardware Design

DGX agent

arXiv:2604.05983v2 Announce Type: replace-cross Abstract: We present Arch (AI-native Register-transfer Clocked Hardware), a hardware description language for micro-architecture specification and AI-as

safetyarxiv-cs-cl
21 Apr 2026
Research

Are Emotion and Rhetoric Neurons in LLM? Neuron Recognition and Adaptive Masking for Emotion-Rhetoric Prediction Steering

DGX agent

arXiv:2604.17255v1 Announce Type: new Abstract: Accurate comprehension and controllable generation of emotion and rhetoric are pivotal for enhancing the reasoning capabilities of large language models

researcharxiv-cs-cl
21 Apr 2026
Applications

Are they lovers or friends? Evaluating LLMs' Social Reasoning in English and Korean Dialogues

DGX agent

arXiv:2510.19028v3 Announce Type: replace Abstract: As LLMs are increasingly deployed in real-world interactions, their social reasoning in interpersonal communication becomes critical. To explore the

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

ArgBench: Benchmarking LLMs on Computational Argumentation Tasks

DGX agent

arXiv:2604.17366v1 Announce Type: new Abstract: Argumentation skills are an essential toolkit for large language models (LLMs). These skills are crucial in various use cases, including self-reflection

model-releasesarxiv-cs-cl
21 Apr 2026
Tutorials

Argument Reconstruction as Supervision for Critical Thinking in LLMs

DGX agent

arXiv:2603.17432v2 Announce Type: replace Abstract: To think critically about arguments, human learners are trained to identify, reconstruct, and evaluate arguments. Argument reconstruction is especia

tutorialsarxiv-cs-cl
21 Apr 2026
Model Releases

ATLAS: Constitution-Conditioned Latent Geometry and Redistribution Across Language Models and Neural Perturbation Data

DGX agent

arXiv:2604.17663v1 Announce Type: cross Abstract: Constitution-conditioned post-training can be analysed as a structured perturbation of a model's learned representational geometry. We introduce ATLAS

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Audio-DeepThinker: Progressive Reasoning-Aware Reinforcement Learning for High-Quality Chain-of-Thought Emergence in Audio Language Models

DGX agent

arXiv:2604.18187v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) have made significant progress in audio understanding, yet they primarily operate as perception-and-answer systems

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Auditing Support Strategies in LLMs through Grounded Multi-Turn Social Simulation

DGX agent

arXiv:2604.17079v1 Announce Type: new Abstract: When users seek social support from chatbots, they disclose their situation gradually, yet most evaluations of supportive LLMs rely on single-turn, full

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

AutoGraph-R1: End-to-End Reinforcement Learning for Knowledge Graph Construction

DGX agent

arXiv:2510.15339v3 Announce Type: replace Abstract: Building effective knowledge graphs (KGs) for Retrieval-Augmented Generation (RAG) is pivotal for advancing question answering (QA) systems. However

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Automatic Slide Updating with User-Defined Dynamic Templates and Natural Language Instructions

DGX agent

arXiv:2604.17894v1 Announce Type: new Abstract: Presentation slides are a primary medium for data-driven reporting, yet keeping complex, analytics-style decks up to date remains labor-intensive. Exist

model-releasesarxiv-cs-cl
21 Apr 2026
Applications

Automatic Speech Recognition for Documenting Endangered Languages: Case Study of Ikema Miyakoan

DGX agent

arXiv:2603.26248v2 Announce Type: replace Abstract: Language endangerment poses a major challenge to linguistic diversity worldwide, and technological advances have opened new avenues for documentatio

applicationsarxiv-cs-cl
21 Apr 2026
Research

AutoRubric: Rubric-Based Generative Rewards for Faithful Multimodal Reasoning

DGX agent

arXiv:2510.14738v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have rapidly advanced from perception tasks to complex multi-step reasoning, yet reinforcement learning wit

researcharxiv-cs-cl
21 Apr 2026
Applications

BASIL: Bayesian Assessment of Sycophancy in LLMs

DGX agent

arXiv:2508.16846v5 Announce Type: replace-cross Abstract: Sycophancy (overly agreeable or flattering behavior) poses a fundamental challenge for human-AI collaboration, particularly in high-stakes dec

applicationsarxiv-cs-cl
21 Apr 2026
Research

Before You Interpret the Profile: Validity Scaling for LLM Metacognitive Self-Report

DGX agent

arXiv:2604.17707v1 Announce Type: new Abstract: Clinical personality assessment screens response validity before interpreting substantive scales. LLM evaluation does not. We apply the validity scaling

researcharxiv-cs-cl
21 Apr 2026
Model Releases

BEFT: Bias-Efficient Fine-Tuning of Language Models in Low-Data Regimes

DGX agent

arXiv:2509.15974v2 Announce Type: replace Abstract: Fine-tuning the bias terms of large language models (LLMs) has the potential to achieve unprecedented parameter efficiency while maintaining competi

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

BenchMarker: An Education-Inspired Toolkit for Highlighting Flaws in Multiple-Choice Benchmarks

DGX agent

arXiv:2602.06221v2 Announce Type: replace Abstract: Multiple-choice question answering (MCQA) is standard in NLP, but benchmarks lack rigorous quality control. We present BenchMarker, an education-ins

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Benchmarking Real-Time Question Answering via Executable Code Workflows

DGX agent

arXiv:2604.16349v1 Announce Type: cross Abstract: Retrieving real-time information is a fundamental capability for search-integrated agents in real-world applications. However, existing benchmarks are

model-releasesarxiv-cs-cl
21 Apr 2026
← Previous
1…130131132133134…161
Next →