AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Agents

GenericAgent: A Token-Efficient Self-Evolving LLM Agent via Contextual Information Density Maximization (V1.0)

DGX agent

arXiv:2604.17091v1 Announce Type: new Abstract: Long-horizon large language model (LLM) agents are fundamentally limited by context. As interactions become longer, tool descriptions, retrieved memorie

agentsarxiv-cs-cl
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Geometric Stability: The Missing Axis of Representations

DGX agent

arXiv:2601.09173v4 Announce Type: replace-cross Abstract: Representational similarity analysis and related methods have become standard tools for comparing the internal geometries of neural networks a

safetyarxiv-cs-cl
21 Apr 2026
Safety

GeometryZero: Advancing Geometry Solving via Group Contrastive Policy Optimization

DGX agent

arXiv:2506.07160v3 Announce Type: replace Abstract: Recent progress in large language models (LLMs) has boosted mathematical reasoning, yet geometry remains challenging where auxiliary construction is

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

GeoRC: A Benchmark for Geolocation Reasoning Chains

DGX agent

arXiv:2601.21278v2 Announce Type: replace-cross Abstract: Vision Language Models (VLMs) are good at recognizing the global location of a photograph -- their geolocation prediction accuracy rivals the

model-releasesarxiv-cs-cl
21 Apr 2026
Research

GoCoMA: Hyperbolic Multimodal Representation Fusion for Large Language Model-Generated Code Attribution

DGX agent

arXiv:2604.16377v1 Announce Type: new Abstract: Large Language Models (LLMs) trained on massive code corpora are now increasingly capable of generating code that is hard to distinguish from human-writ

researcharxiv-cs-cl
21 Apr 2026
Local Ai

GraSP: Graph-Structured Skill Compositions for LLM Agents

DGX agent

arXiv:2604.17870v1 Announce Type: new Abstract: Skill ecosystems for LLM agents have matured rapidly, yet recent benchmarks show that providing agents with more skills does not monotonically improve p

local-aiarxiv-cs-cl
21 Apr 2026
Model Releases

GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling

DGX agent

arXiv:2604.18556v1 Announce Type: new Abstract: Weight quantization has become a standard tool for efficient LLM deployment, especially for local inference, where models are now routinely served at 2-

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

HalluSAE: Detecting Hallucinations in Large Language Models via Sparse Auto-Encoders

DGX agent

arXiv:2604.16430v1 Announce Type: new Abstract: Large Language Models (LLMs) are powerful and widely adopted, but their practical impact is limited by the well-known hallucination phenomenon. While re

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Hard to Be Heard: Phoneme-Level ASR Analysis of Phonologically Complex, Low-Resource Endangered Languages

DGX agent

arXiv:2604.18204v1 Announce Type: new Abstract: We present a phoneme-level analysis of automatic speech recognition (ASR) for two low-resourced and phonologically complex East Caucasian languages, Arc

researcharxiv-cs-cl
21 Apr 2026
Agents

HeLa-Mem: Hebbian Learning and Associative Memory for LLM Agents

DGX agent

arXiv:2604.16839v1 Announce Type: new Abstract: Long-term memory is a critical challenge for Large Language Model agents, as fixed context windows cannot preserve coherence across extended interaction

agentsarxiv-cs-cl
21 Apr 2026
Research

HeteroCache: A Dynamic Retrieval Approach to Heterogeneous KV Cache Compression for Long-Context LLM Inference

DGX agent

arXiv:2601.13684v2 Announce Type: replace Abstract: The linear memory growth of the KV cache poses a significant bottleneck for LLM inference in long-context tasks. Existing static compression methods

researcharxiv-cs-cl
21 Apr 2026
Tutorials

Heterogeneity in Formal Linguistic Competence of Language Models: Is Data the Real Bottleneck?

DGX agent

arXiv:2604.17930v1 Announce Type: new Abstract: Large Language Models (LLMs) exhibit a puzzling disparity in their formal linguistic competence: while they learn some linguistic phenomena with near-pe

tutorialsarxiv-cs-cl
21 Apr 2026
Applications

Hierarchical Retrieval with Out-Of-Vocabulary Queries: A Case Study on SNOMED CT

DGX agent

arXiv:2511.16698v2 Announce Type: replace Abstract: SNOMED CT is a biomedical ontology with a hierarchical representation, modelling terminological concepts at a large scale. Knowledge retrieval in SN

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

HiGMem: A Hierarchical and LLM-Guided Memory System for Long-Term Conversational Agents

DGX agent

arXiv:2604.18349v1 Announce Type: new Abstract: Long-term conversational large language model (LLM) agents require memory systems that can recover relevant evidence from historical interactions withou

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

HiP-LoRA: Budgeted Spectral Plasticity for Robust Low-Rank Adaptation

DGX agent

arXiv:2604.17751v1 Announce Type: cross Abstract: Adapting foundation models under resource budgets relies heavily on Parameter-Efficient Fine-Tuning (PEFT), with LoRA being a standard modular solutio

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

HiRAS: A Hierarchical Multi-Agent Framework for Paper-to-Code Generation and Execution

DGX agent

arXiv:2604.17745v1 Announce Type: new Abstract: Recent advances in large language models have highlighted their potential to automate computational research, particularly reproducing experimental resu

model-releasesarxiv-cs-cl
21 Apr 2026
Applications

HopRank: Self-Supervised LLM Preference-Tuning on Graphs for Few-Shot Node Classification

DGX agent

arXiv:2604.17271v1 Announce Type: new Abstract: Node classification on text-attributed graphs (TAGs) is a fundamental task with broad applications in citation analysis, social networks, and recommenda

applicationsarxiv-cs-cl
21 Apr 2026
Research

HopWeaver: Cross-Document Synthesis of High-Quality and Authentic Multi-Hop Questions

DGX agent

arXiv:2505.15087v3 Announce Type: replace Abstract: Multi-Hop Question Answering (MHQA) is crucial for evaluating the model's capability to integrate information from diverse sources. However, creatin

researcharxiv-cs-cl
21 Apr 2026
Model Releases

HORIZON: A Benchmark for In-the-wild User Behaviour Modeling

DGX agent

arXiv:2604.17259v1 Announce Type: cross Abstract: User behavior in the real world is diverse, cross-domain, and spans long time horizons. Existing user modeling benchmarks however remain narrow, focus

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

HorizonBench: Long-Horizon Personalization with Evolving Preferences

DGX agent

arXiv:2604.17283v1 Announce Type: new Abstract: User preferences evolve across months of interaction, and tracking them requires inferring when a stated preference has been changed by a subsequent lif

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

How Creative Are Large Language Models in Generating Molecules?

DGX agent

arXiv:2604.18031v1 Announce Type: new Abstract: Molecule generation requires satisfying multiple chemical and biological constraints while searching a large and structured chemical space. This makes i

local-aiarxiv-cs-cl
21 Apr 2026
Safety

How Language Models Conflate Logical Validity with Plausibility: A Representational Analysis of Content Effects

DGX agent

arXiv:2510.06700v3 Announce Type: replace Abstract: Both humans and large language models (LLMs) exhibit content effects: biases in which the plausibility of the semantic content of a reasoning proble

safetyarxiv-cs-cl
21 Apr 2026
Applications

How Non-Linguistic Is the Indus Sign System? A Synthetic-Baseline Scorecard

DGX agent

arXiv:2604.17828v1 Announce Type: new Abstract: Whether the Indus Valley sign system (c. 2600-1900 BCE) encodes spoken language has been debated for decades. This paper introduces a multi-metric discr

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study

DGX agent

arXiv:2505.15404v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have achieved remarkable success on reasoning-intensive tasks such as mathematics and programming. However, their enha

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

How Tokenization Limits Phonological Knowledge Representation in Language Models and How to Improve Them

DGX agent

arXiv:2604.17105v1 Announce Type: new Abstract: Tokenization is the first step in every language model (LM), yet it never takes the sounds of words into account. We investigate how tokenization influe

local-aiarxiv-cs-cl
21 Apr 2026
Applications

How Training Data Shapes the Use of Parametric and In-Context Knowledge in Language Models

DGX agent

arXiv:2510.02370v3 Announce Type: replace Abstract: Large language models leverage both parametric knowledge acquired during pretraining and in-context knowledge provided at inference time. Crucially,

applicationsarxiv-cs-cl
21 Apr 2026
Research

HPLT 3.0: Very Large-Scale Multilingual Resources for LLMs and MT. Mono- and Bi-lingual Data, Multilingual Evaluation, and Pre-Trained Models

DGX agent

arXiv:2511.01066v3 Announce Type: replace Abstract: We present an ongoing initiative to provide open, very large, high-quality, and richly annotated textual datasets for almost 200 languages. At 30 tr

researcharxiv-cs-cl
21 Apr 2026
Safety

Human-Centered Supervision for Sentiment Analysis in Telugu: A Systematic Inquiry Beyond Accuracy

DGX agent

arXiv:2508.01486v3 Announce Type: replace Abstract: Sentiment analysis for low-resource languages remains challenging in an era where interpretability, human alignment, and fairness are increasingly n

safetyarxiv-cs-cl
21 Apr 2026
Safety

IceBreaker for Conversational Agents: Breaking the First-Message Barrier with Personalized Starters

DGX agent

arXiv:2604.18375v1 Announce Type: new Abstract: Conversational agents, such as ChatGPT and Doubao, have become essential daily assistants for billions of users. To further enhance engagement, these sy

safetyarxiv-cs-cl
21 Apr 2026
Research

ICLAD: In-Context Learning with Comparison-Guidance for Audio Deepfake Detection

DGX agent

arXiv:2604.16749v1 Announce Type: cross Abstract: Audio deepfakes pose a significant security threat, yet current state-of-the-art (SOTA) detection systems do not generalize well to realistic in-the-w

researcharxiv-cs-cl
21 Apr 2026
Research

ImpRIF: Stronger Implicit Reasoning Leads to Better Complex Instruction Following

DGX agent

arXiv:2602.21228v2 Announce Type: replace Abstract: As applications of large language models (LLMs) become increasingly complex, the demand for robust complex instruction following capabilities is gro

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Improving LLM Code Reasoning via Semantic Equivalence Self-Play with Formal Verification

DGX agent

arXiv:2604.17010v1 Announce Type: new Abstract: We introduce a self-play framework for semantic equivalence in Haskell, utilizing formal verification to guide adversarial training between a generator

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Improving Speech Recognition of Named Entities in Classroom Speech with LLM Revision and Phonetic-Semantic Context

DGX agent

arXiv:2506.10779v2 Announce Type: replace Abstract: Classroom speech and lectures often contain named entities (NEs) such as names of people and special terminology. While automatic speech recognition

researcharxiv-cs-cl
21 Apr 2026
Research

Incentivizing Parametric Knowledge via Reinforcement Learning with Verifiable Rewards for Cross-Cultural Entity Translation

DGX agent

arXiv:2604.16881v1 Announce Type: new Abstract: Cross-cultural entity translation remains challenging for large language models (LLMs) as literal or phonetic renderings are usually yielded instead of

researcharxiv-cs-cl
21 Apr 2026
Safety

Inertia in Moral and Value Judgments of Large Language Models

DGX agent

arXiv:2408.09049v3 Announce Type: replace Abstract: Large Language Models (LLMs) behave non-deterministically, and prompting has become a common method for steering their outputs. A popular strategy i

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Inflated Excellence or True Performance? Rethinking Medical Diagnostic Benchmarks with Dynamic Evaluation

DGX agent

arXiv:2510.09275v2 Announce Type: replace Abstract: Medical diagnostics is a high-stakes and complex domain that is critical to patient care. However, current evaluations of large language models (LLM

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Information Representation Fairness in Long-Document Embeddings: The Peculiar Interaction of Positional and Language Bias

DGX agent

arXiv:2601.16934v2 Announce Type: replace Abstract: To be discoverable in an embedding-based search process, each part of a document should be reflected in its embedding representation. To quantify an

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Injecting Structured Biomedical Knowledge into Language Models: Continual Pretraining vs. GraphRAG

DGX agent

arXiv:2604.16422v1 Announce Type: new Abstract: The injection of domain-specific knowledge is crucial for adapting language models (LMs) to specialized fields such as biomedicine. While most current a

model-releasesarxiv-cs-cl
21 Apr 2026
Research

iPhoneme: Brain-to-Text Communication for ALS Using ConformerXL Decoding

DGX agent

arXiv:2604.16441v1 Announce Type: cross Abstract: Brain-computer interfaces (BCIs) for speech restoration hold transformative potential for the approximately 173,000--232,500 individuals worldwide wit

researcharxiv-cs-cl
21 Apr 2026
Agents

Is Agentic RAG worth it? An experimental comparison of RAG approaches

DGX agent

arXiv:2601.07711v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems are usually defined by the combination of a generator and a retrieval component that extracts textual c

agentsarxiv-cs-cl
21 Apr 2026
Safety

IYKYK (But AI Doesn't): Automated Content Moderation Does Not Capture Communities' Heterogeneous Attitudes Towards Reclaimed Language

DGX agent

arXiv:2604.16654v1 Announce Type: new Abstract: Reclaimed slur usage is a common and meaningful practice online for many marginalized communities. It serves as a source of solidarity, identity, and sh

safetyarxiv-cs-cl
21 Apr 2026
Safety

Jailbreaking Large Language Models with Morality Attacks

DGX agent

arXiv:2604.17053v1 Announce Type: new Abstract: Pluralism alignment with AI has the sophisticated and necessary goal of creating AI that can coexist with and serve morally multifaceted humanity. Resea

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

JudgeMeNot: Personalizing Large Language Models to Emulate Judicial Reasoning in Hebrew

DGX agent

arXiv:2604.18041v1 Announce Type: new Abstract: Despite significant advances in large language models, personalizing them for individual decision-makers remains an open problem. Here, we introduce a s

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Jupiter-N Technical Report

DGX agent

arXiv:2604.17429v1 Announce Type: new Abstract: We present Jupiter-N, a hybrid reasoning model post-trained from Nemotron 3 Super, a fully open-source 120 billion parameter LLM. We target three object

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Knowing When to Quit: A Principled Framework for Dynamic Abstention in LLM Reasoning

DGX agent

arXiv:2604.18419v1 Announce Type: cross Abstract: Large language models (LLMs) using chain-of-thought reasoning often waste substantial compute by producing long, incorrect responses. Abstention can m

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Language Models Don't Know What You Want: Evaluating Personalization in Deep Research Needs Real Users

DGX agent

arXiv:2603.16120v2 Announce Type: replace Abstract: Deep Research (DR) systems help researchers cope with ballooning publishing counts. Such tools synthesize scientific papers to answer research queri

model-releasesarxiv-cs-cl
21 Apr 2026
Applications

Large Language Models Are Bad Dice Players: LLMs Struggle to Generate Random Numbers from Statistical Distributions

DGX agent

arXiv:2601.05414v2 Announce Type: replace Abstract: As large language models (LLMs) transition from chat interfaces to integral components of stochastic pipelines and systems approaching general intel

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

Large Language Models Are Still Misled by Simple Bias Ensembles

DGX agent

arXiv:2505.16522v3 Announce Type: replace Abstract: With the evolution of large language models (LLMs), their robustness against individual simple biases has been enhanced. However, we observe that th

model-releasesarxiv-cs-cl
21 Apr 2026
← Previous
1…134135136137138…161
Next →