AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
Model Releases

Multilingual Steering by Design: Multilingual Sparse Autoencoders and Principled Layer Selection

DGX agent

arXiv:2605.23036v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) enable feature-level mechanistic interpretability and activation steering in large language models (LLMs), but SAE-based lang

model-releasesarxiv-cs-cl
25 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Naturalistic measure of social norms alignment

DGX agent

arXiv:2605.23420v1 Announce Type: new Abstract: Social norms reflect shared expectations on acceptable behavior. Measuring social norms alignment remains challenging, with existing approaches typicall

safetyarxiv-cs-cl
25 May 2026
Safety

NLG Evaluation: Past, Present, Future

DGX agent

arXiv:2605.23715v1 Announce Type: new Abstract: Natural Language Generation (NLG) evaluation has changed dramatically since 1990, and will continue to evolve in the future. In 1990, when NLG had close

safetyarxiv-cs-cl
25 May 2026
Model Releases

OpenSkillEval: Automatically Auditing the Open Skill Ecosystem for LLM Agents

DGX agent

arXiv:2605.23657v1 Announce Type: new Abstract: Skills, i.e., structured workflow instructions distilled for large language models (LLMs), are becoming an increasingly important mechanism for improvin

model-releasesarxiv-cs-cl
25 May 2026
Research

Pooling and Semantic Shift: The Fundamental Challenges in Long Text Embedding and Retrieval

DGX agent

arXiv:2603.21437v2 Announce Type: replace Abstract: Transformer-based embedding models frequently exhibit geometric pathologies, such as anisotropy and length-induced representation collapse, which ca

researcharxiv-cs-cl
25 May 2026
Model Releases

PROGRESSLM: Towards Progress Reasoning in Vision-Language Models

DGX agent

arXiv:2601.15224v2 Announce Type: replace-cross Abstract: Estimating task progress requires reasoning over long-horizon dynamics rather than recognizing static visual content. While modern Vision-Lang

model-releasesarxiv-cs-cl
25 May 2026
Agents

Query-Adaptive Semantic Chunking for Retrieval-Augmented Generation: A Dynamic Strategy with Contextual Window Expansion

DGX agent

arXiv:2605.22834v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems depend critically on document chunking quality for retrieving relevant context. Fixed chunking segments doc

agentsarxiv-cs-cl
25 May 2026
Research

RADAR: Relative Angular Divergence Across Representations

DGX agent

arXiv:2605.23028v1 Announce Type: cross Abstract: Machine learning methods rely on data. However, gathering suitable data can be challenging due to availability constraints, cost, or the need for doma

researcharxiv-cs-cl
25 May 2026
Research

RAS: Reflection-Augmented Scaling with In-Context Learning for Executable Cypher Query Generation

DGX agent

arXiv:2605.22937v1 Announce Type: new Abstract: Inference-time scaling can reduce errors in structured query generation, but methods to allocate the compute for query code generation remains underexpl

researcharxiv-cs-cl
25 May 2026
Research

Robust LLM Watermarking with Minimal Semantic Distortion for IP Protection

DGX agent

arXiv:2605.23175v1 Announce Type: cross Abstract: Proprietary large language models (LLMs) face risks of intellectual property (IP) violation, as adversaries can replicate an LLM by collecting input-o

researcharxiv-cs-cl
25 May 2026
Model Releases

Same Model, Different Weakness: How Language and Modality Reshape the Jailbreak Attack Surface in Frontier MLLMs

DGX agent

arXiv:2605.23157v1 Announce Type: new Abstract: The attack surface of a multimodal large language model (MLLM) is language-dependent in ways that reveal the mechanistic structure of alignment failures

model-releasesarxiv-cs-cl
25 May 2026
Research

SciNet: Evaluating AI Agents in Relation-Aware Scientific Literature Retrieval

DGX agent

arXiv:2601.03260v2 Announce Type: replace-cross Abstract: AI agents have seen widespread adoption in information retrieval for scientific research, giving rise to tools such as Deep Research. However,

researcharxiv-cs-cl
25 May 2026
Research

Self-Improving In-Context Learning

DGX agent

arXiv:2605.23180v1 Announce Type: new Abstract: We propose to improve in-context learning (ICL) by optimizing the continuous embeddings of a fixed few-shot prompt at test time. The key observation is

researcharxiv-cs-cl
25 May 2026
Model Releases

SemEval-2026 Task 6: CLARITY -- Unmasking Political Question Evasions

DGX agent

arXiv:2603.14027v2 Announce Type: replace Abstract: Political speakers often avoid answering questions directly while maintaining the appearance of responsiveness. Despite its importance for public di

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Speak-to-Structure: Evaluating LLMs in Open-domain Natural Language-Driven Molecule Generation

DGX agent

arXiv:2412.14642v4 Announce Type: replace Abstract: Recently, Large Language Models (LLMs) have demonstrated great potential in natural language-driven molecule discovery. However, existing datasets a

model-releasesarxiv-cs-cl
25 May 2026
Research

Strong Teacher Not Needed? On Distillation in LLM Pretraining

DGX agent

arXiv:2605.23857v1 Announce Type: cross Abstract: Knowledge distillation generally assumes a strong-to-weak relationship where stronger teachers yield better students. In this work, we examine this as

researcharxiv-cs-cl
25 May 2026
Applications

Structure-Guided Entity Resolution: Fine-Tuning LLMs for Robust Name Matching in Complex Linguistic Contexts

DGX agent

arXiv:2605.23597v1 Announce Type: new Abstract: Matching person names across heterogeneous records is a core challenge in entity resolution, especially within linguistically and culturally complex env

applicationsarxiv-cs-cl
25 May 2026
Model Releases

TEAM: Temporal-Spatial Consistency Guided Expert Activation for MoE Diffusion Language Model Acceleration

DGX agent

arXiv:2602.08404v2 Announce Type: replace Abstract: Diffusion large language models (dLLMs) have recently gained significant attention due to their inherent support for parallel decoding. Building on

model-releasesarxiv-cs-cl
25 May 2026
Research

The Efficiency Frontier: A Unified Framework for Cost-Performance Optimization in LLM Context Management

DGX agent

arXiv:2605.23071v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on long-context processing, but expanding context windows introduces substantial computational and financ

researcharxiv-cs-cl
25 May 2026
Research

TurkicNLP: An NLP Toolkit for Turkic Languages

DGX agent

arXiv:2602.19174v5 Announce Type: replace Abstract: Natural language processing for the Turkic language family, spoken by over 200 million people across Eurasia, remains fragmented, with most language

researcharxiv-cs-cl
25 May 2026
Model Releases

Vector Retrieval with Similarity and Diversity: How Hard Is It?

DGX agent

arXiv:2407.04573v4 Announce Type: replace-cross Abstract: Dense vector retrieval is an important building block of modern machine learning systems, underlying applications ranging from semantic search

model-releasesarxiv-cs-cl
25 May 2026
Research

What Does the Server See? Understanding Privacy Leakage from Large Language Models in Split Inference

DGX agent

arXiv:2605.23158v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) on resource-constrained devices remains challenging, spurring interest in split inference, where models

researcharxiv-cs-cl
25 May 2026
Model Releases

What Training Data Teaches RL Memory Agents: An Empirical Study of Curriculum Effects in Memory-Augmented QA

DGX agent

arXiv:2605.23067v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a viable recipe for training LLM agents to reason over external memory banks in multi-session dialogue. Exist

model-releasesarxiv-cs-cl
25 May 2026
Applications

When AI Takes Sides on Questions of Faith: Persistent Asymmetries in AI-Mediated Faith Guidance

DGX agent

arXiv:2605.22975v1 Announce Type: new Abstract: We ask whether large language models (LLMs) treat queries about religious conversion symmetrically. The answer is no. When asked for advice on hypotheti

applicationsarxiv-cs-cl
25 May 2026
Agents

When Is Next-Token Prediction Useful? Marginalization, Ergodicity, Mixture Identifiability, Local Sufficiency, RAG, Tools, and Programming

DGX agent

arXiv:2605.23278v1 Announce Type: new Abstract: Language models trained on observed sequences are often described as learning the conditional distribution of the next token given previous tokens. This

agentsarxiv-cs-cl
25 May 2026
Model Releases

When Symptoms Are Not Enough: Evidence-Weighting Patterns in Large Language Model Psychiatric Screening

DGX agent

arXiv:2605.23148v1 Announce Type: new Abstract: As demand for mental health care outpaces clinician-delivered assessment, scalable screening tools are increasingly needed. Large language models (LLMs)

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

A Comparative Study of Language Models for Khmer Retrieval-Augmented Question Answering

DGX agent

arXiv:2605.22099v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has emerged as a promising paradigm for grounding large language model (LLM) outputs in retrieved evidence, thereby

model-releasesarxiv-cs-cl
22 May 2026
Tutorials

A Tutorial on Diffusion Theory: From Differential Equations to Diffusion Models

DGX agent

arXiv:2605.22586v1 Announce Type: cross Abstract: This tutorial develops diffusion models from the viewpoint of differential equations. We begin with the conditional Gaussian forward process and show

tutorialsarxiv-cs-cl
22 May 2026
Agents

ACC: Compiling Agent Trajectories for Long-Context Training

DGX agent

arXiv:2605.21850v1 Announce Type: new Abstract: Recent development of agents has renewed demand for long-context reasoning capacity of LLMs. However, training LLMs for this capacity requires costly lo

agentsarxiv-cs-cl
22 May 2026
Research

Accelerated Test-Time Scaling with Model-Free Speculative Sampling

DGX agent

arXiv:2506.04708v3 Announce Type: replace Abstract: Language models have demonstrated remarkable capabilities in reasoning tasks through test-time scaling techniques like best-of-N sampling and tree s

researcharxiv-cs-cl
22 May 2026
Safety

Agentic CLEAR: Automating Multi-Level Evaluation of LLM Agents

DGX agent

arXiv:2605.22608v1 Announce Type: new Abstract: Agentic systems are becoming more capable: agents define strategies, take actions, and interact with different environments. This autonomy poses serious

safetyarxiv-cs-cl
22 May 2026
Model Releases

AMEL: Accumulated Message Effects on LLM Judgments

DGX agent

arXiv:2605.22714v1 Announce Type: cross Abstract: Large language models are routinely used as automated evaluators: to review code, moderate content, or score outputs, often with many items passing th

model-releasesarxiv-cs-cl
22 May 2026
Safety

Amplifying, Not Learning: Fine-Tuned AI Text Detectors Amplify a Pretrained Direction

DGX agent

arXiv:2605.21653v1 Announce Type: cross Abstract: AI text detectors amplify a pretrained typicality axis; they do not construct an AI-vs-human boundary. On raw encoders before any task supervision, pr

safetyarxiv-cs-cl
22 May 2026
Agents

An Entity Linking Agent for Question Answering

DGX agent

arXiv:2508.03865v4 Announce Type: replace Abstract: Some Question Answering (QA) systems rely on knowledge bases (KBs) to provide accurate answers. Entity Linking (EL) plays a critical role in linking

agentsarxiv-cs-cl
22 May 2026
Tutorials

AnyMo: Geometry-Aware Setup-Agnostic Modeling of Human Motion in the Wild

DGX agent

arXiv:2605.22715v1 Announce Type: cross Abstract: As wearable and mobile devices become increasingly embedded in daily life, they offer a practical way to continuously sense human motion in the wild.

tutorialsarxiv-cs-cl
22 May 2026
Model Releases

ArabDiscrim: A Decade-Long Arabic Facebook Corpus on Racism and Discrimination

DGX agent

arXiv:2605.22081v1 Announce Type: new Abstract: We present ArabDiscrim, a decade-long lexical resource and corpus of 293K public Arabic Facebook posts (2014--2024) discussing racism and discrimination

model-releasesarxiv-cs-cl
22 May 2026
Research

Assisted Counterspeech Writing at the Crossroads of Hate Speech and Misinformation

DGX agent

arXiv:2605.22435v1 Announce Type: new Abstract: Hate speech and misinformation frequently co-occur online, amplifying prejudice and polarization. Given their scale, using Large Language Models (LLMs)

researcharxiv-cs-cl
22 May 2026
Model Releases

Audience Engagement with Arabic Women's Social Empowerment and Wellbeing: A Decadal Corpus

DGX agent

arXiv:2605.22204v1 Announce Type: new Abstract: This paper presents the Arabic Women and Society Corpus, a ten year collection of 252,487 public Arabic Facebook posts related to women's empowerment an

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

BEiTScore: Reference-free Image Captioning Evaluation with an Efficient Cross-Encoder Model

DGX agent

arXiv:2605.21728v1 Announce Type: cross Abstract: Image captioning evaluation remains a significant challenge, as vision-language models evolve toward more challenging capabilities such as generating

model-releasesarxiv-cs-cl
22 May 2026
Applications

BeLink: Biomedical Entity Linking Meets Generative Re-Ranking

DGX agent

arXiv:2605.22501v1 Announce Type: new Abstract: Despite recent progress, Biomedical Entity Linking (BEL) with large language models (LLMs) remains computationally inefficient and challenging to deploy

applicationsarxiv-cs-cl
22 May 2026
Model Releases

Beyond Acoustic Emotion Recognition: Multimodal Pathos Analysis in Political Speech Using LLM-Based and Acoustic Emotion Models

DGX agent

arXiv:2605.22732v1 Announce Type: cross Abstract: We investigate whether acoustic emotion recognition models can serve as proxies for the Pathos dimension in political speech analysis, as operationali

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Beyond Benchmark Islands: Toward Representative Trustworthiness Evaluation for Agentic AI

DGX agent

arXiv:2603.14987v2 Announce Type: replace Abstract: Agentic AI systems increasingly act through tool-augmented, multi-step workflows whose failures (unsafe tool use, unauthorised actions, social harm)

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Beyond Temperature: Hyperfitting as a Late-Stage Geometric Expansion

DGX agent

arXiv:2605.22579v1 Announce Type: new Abstract: Recent work has identified a counterintuitive phenomenon termed 'Hyperfitting', where fine-tuning Large Language Models (LLMs) to near-zero training los

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Blind Spots in the Guard: How Domain-Camouflaged Injection Attacks Evade Detection in Multi-Agent LLM Systems

DGX agent

arXiv:2605.22001v1 Announce Type: cross Abstract: Injection detectors deployed to protect LLM agents are calibrated on static, template-based payloads that announce themselves as override directives.

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety

DGX agent

arXiv:2605.22643v1 Announce Type: new Abstract: Background. Traditional safety benchmarks for language models evaluate generated text: whether a model outputs toxic language, reproduces bias, or follo

model-releasesarxiv-cs-cl
22 May 2026
Local Ai

Boundary-targeted Membership Inference Attacks on Safety Classifiers

DGX agent

arXiv:2605.22373v1 Announce Type: cross Abstract: Safety classifiers are essential safeguards within generative AI systems, filtering harmful content or identifying at-risk users when interacting with

local-aiarxiv-cs-cl
22 May 2026
Local Ai

Broadening Access to Transportation Safety Data with Generative AI: A Schema-Grounded Framework for Spatial Natural Language Queries

DGX agent

arXiv:2605.21712v1 Announce Type: new Abstract: Transportation safety analysis requires integrating crash records, roadway attributes, and geospatial data through GIS-based workflows, but access remai

local-aiarxiv-cs-cl
22 May 2026
Model Releases

Check Your LLM's Secret Dictionary! Five Lines of Code Reveal What Your LLM Learned (Including What It Shouldn't Have)

DGX agent

arXiv:2605.22005v1 Announce Type: cross Abstract: We show that singular value decomposition of the lm_head} weight matrix of a transformer-based large language model -- requiring only five lines of Py

model-releasesarxiv-cs-cl
22 May 2026
← Previous
1…8081828384…162
Next →