AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
Model Releases

Fix the Structural Bottleneck: Context Compression via Explicit Information Transmission

DGX agent

arXiv:2602.03784v2 Announce Type: replace Abstract: Long-context LLM agents often struggle with growing token, memory, and latency costs, making efficient context compression essential for practical d

model-releasesarxiv-cs-cl
19 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

FOL2NS: Generating Natural Sentences from First-Order Logic

DGX agent

arXiv:2605.18155v1 Announce Type: new Abstract: Translating formal language into natural language is a foundational challenge in NLP, driving various downstream applications in semantic parsing, theor

researcharxiv-cs-cl
19 May 2026
Research

Forecasting Downstream Performance of LLMs With Proxy Metrics

DGX agent

arXiv:2605.18607v1 Announce Type: new Abstract: Progress in language model development is often driven by comparative decisions: which architecture to adopt, which pretraining corpus to use, or which

researcharxiv-cs-cl
19 May 2026
Research

From BERT to T5: A Study of Named Entity Recognition

DGX agent

arXiv:2605.18462v1 Announce Type: new Abstract: Named entity recognition (NER) has been one of the essential preliminary steps in modern NLP applications. This report focuses on implementing the NER t

researcharxiv-cs-cl
19 May 2026
Applications

From Documents to Segments: A Contextual Reformulation for Topic Assignment

DGX agent

arXiv:2605.17714v1 Announce Type: new Abstract: Traditional topic modeling assigns a single topic to each document. In practice, however, many real-world documents, such as product reviews or open-end

applicationsarxiv-cs-cl
19 May 2026
Research

From Isolated Scoring to Collaborative Ranking: A Comparison-Native Framework for LLM-Based Paper Evaluation

DGX agent

arXiv:2603.17588v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are currently applied to scientific paper evaluation by assigning an absolute score to each paper independently.

researcharxiv-cs-cl
19 May 2026
Model Releases

Gated KalmaNet: A Fading Memory Layer Through Test-Time Ridge Regression

DGX agent

arXiv:2511.21016v3 Announce Type: replace-cross Abstract: Linear State-Space Models (SSMs) offer an efficient alternative to softmax Attention with constant memory and linear compute, but their lossy,

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

General Preference Reinforcement Learning

DGX agent

arXiv:2605.18721v1 Announce Type: cross Abstract: Post-training has split large language model (LLM) alignment into two largely disconnected tracks. Online reinforcement learning (RL) with verifiable

model-releasesarxiv-cs-cl
19 May 2026
Agents

Generative AI Advertising as a Problem of Trustworthy Commercial Intervention

DGX agent

arXiv:2605.18673v1 Announce Type: cross Abstract: Major deployed generative AI advertising systems preserve a visible boundary between commercial content and AI-generated responses. Yet empirical rese

agentsarxiv-cs-cl
19 May 2026
Model Releases

Generative Artificial Intelligence for Literature Reviews

DGX agent

arXiv:2605.16475v1 Announce Type: cross Abstract: Generative artificial intelligence (GenAI), based on large-language models (LLMs), such as ChatGPT, has taken organizations, academia, and the public

model-releasesarxiv-cs-cl
19 May 2026
Research

GUT-IS: A Data-Driven Approach to Integrating Constructs and Their Relations in Information Systems

DGX agent

arXiv:2605.18567v1 Announce Type: new Abstract: Structural equation modeling is widely used in IS research. However, inconsistent construct definitions impede the cumulative development of knowledge.

researcharxiv-cs-cl
19 May 2026
Model Releases

HalluScore: Large Language Model Hallucination Question Answering Benchmark

DGX agent

arXiv:2605.17007v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable progress in natural language generation, but remain susceptible to hallucination. In response to g

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

HEED: Density-Weighted Residual Alignment for Hybrid Vision-Language Model Distillation

DGX agent

arXiv:2605.17093v1 Announce Type: cross Abstract: Distilling vision-language models into faster hybrid architectures, such as 3:1 Mamba-2/attention mixes, is now standard practice for making inference

model-releasesarxiv-cs-cl
19 May 2026
Safety

Helpful to a Fault: Measuring Illicit Assistance in Multi-Turn, Multilingual LLM Agents

DGX agent

arXiv:2602.16346v3 Announce Type: replace Abstract: LLM-based agents execute real-world workflows via tools and memory. These affordances enable ill-intended adversaries to also use these agents to ca

safetyarxiv-cs-cl
19 May 2026
Model Releases

How Good LLMs Are at Answering Bangla Medical Visual Questions? Dataset and Benchmarking

DGX agent

arXiv:2605.18111v1 Announce Type: new Abstract: Recent advancements in Large Language Models (LLMs) and Large Vision Language Models (LVLMs) have enabled general-purpose systems to demonstrate promisi

model-releasesarxiv-cs-cl
19 May 2026
Safety

How Loud Rumbles Hit Newsstands: A Data Analysis of Coverage and Spatial Bias in German News about Landslides Around the World

DGX agent

arXiv:2605.18105v1 Announce Type: new Abstract: Landslides often hit newsstands due to their destructive and potentially fatal effects. News are a valuable source of information for creating or enrich

safetyarxiv-cs-cl
19 May 2026
Safety

How Off-Policy Can GRPO Be? Mu-GRPO for Efficient LLM Reinforcement Learning

DGX agent

arXiv:2605.17570v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) has been a key driver of recent progress in reinforcement learning with verifiable rewards (RLVR) for large

safetyarxiv-cs-cl
19 May 2026
Research

Hybrid Feature Combinations with CNN for Bangla Fake News Classification

DGX agent

arXiv:2605.17481v1 Announce Type: new Abstract: Nowadays, people in Bangladesh frequently rely on the internet and social media for daily news instead of traditional newspapers. However, the spread of

researcharxiv-cs-cl
19 May 2026
Model Releases

HyDRA: Hybrid Dynamic Routing Architecture for Heterogeneous LLM Pools

DGX agent

arXiv:2605.17106v1 Announce Type: new Abstract: Production LLM deployments increasingly maintain heterogeneous model pools spanning order-of-magnitude cost differences. Existing routers make binary st

model-releasesarxiv-cs-cl
19 May 2026
Safety

Implicit Hierarchical GRPO: Decoupling Tool Invocation from Execution for Tool-Integrated Mathematical Reasoning

DGX agent

arXiv:2605.18500v1 Announce Type: new Abstract: Large language models (LLMs) have increasingly leveraged tool invocation to enhance their reasoning capabilities. However, existing approaches typically

safetyarxiv-cs-cl
19 May 2026
Research

Infini-News: Efficiently Queryable Access to 1.3 Billion Processed Common Crawl News Articles

DGX agent

arXiv:2605.18337v1 Announce Type: new Abstract: Large-scale news corpora support a wide range of research in Computational Social Science and NLP, yet access remains constrained: commercial archives i

researcharxiv-cs-cl
19 May 2026
Research

Information-Theoretic Storage Cost in Sentence Comprehension

DGX agent

arXiv:2602.18217v2 Announce Type: replace Abstract: Real-time sentence comprehension imposes a significant load on working memory, as comprehenders must maintain contextual information to anticipate f

researcharxiv-cs-cl
19 May 2026
Model Releases

Internalizing Tool Knowledge in Small Language Models via QLoRA Fine-Tuning

DGX agent

arXiv:2605.17774v1 Announce Type: new Abstract: Large language models are increasingly used as planning components in agentic systems, but current tool-use pipelines often require full tool schemas to

model-releasesarxiv-cs-cl
19 May 2026
Tutorials

iPOE: Interpretable Prompt Optimization via Explanations

DGX agent

arXiv:2605.18113v1 Announce Type: new Abstract: Prompt optimization has often been framed as a discrete search problem to find high-performing and robust instructions for an LLM. However, the search r

tutorialsarxiv-cs-cl
19 May 2026
Research

JSPG: Dynamic Dictionary Filtering via Joint Semantic-Pinyin-Glyph Retrieval for Chinese Contextual ASR

DGX agent

arXiv:2605.16896v1 Announce Type: new Abstract: Contextual Automatic Speech Recognition (ASR) faces challenges with large-scale keyword dictionaries, as excessive irrelevant candidates introduce noise

researcharxiv-cs-cl
19 May 2026
Research

Knowledge-to-Verification: Exploring RLVR for LLMs in Knowledge-Intensive Domains

DGX agent

arXiv:2605.18261v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has demonstrated promising potential to enhance the reasoning capabilities of large language model

researcharxiv-cs-cl
19 May 2026
Hardware

KVDrive: A Holistic Multi-Tier KV Cache Management System for Long-Context LLM Inference

DGX agent

arXiv:2605.18071v1 Announce Type: new Abstract: Supporting long-context LLMs is challenging due to the substantial memory demands of the key-value (KV) cache. Existing offloading systems store the ful

hardwarearxiv-cs-cl
19 May 2026
Research

Language Acquisition Device in Large Language Models

DGX agent

arXiv:2605.16758v1 Announce Type: new Abstract: Large Language Models (LLMs) remain substantially less data-efficient than humans. Pre-pretraining (PPT) on synthetic languages has been proposed to clo

researcharxiv-cs-cl
19 May 2026
Model Releases

Language-Switching Triggers Take a Latent Detour Through Language Models

DGX agent

arXiv:2605.18646v1 Announce Type: new Abstract: Backdoor attacks on language models pose a growing security concern, yet the internal mechanisms by which a trigger sequence hijacks model computations

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

LaPA^2: Length-Aware Prefix and Prompt Attention Augmentation for Long-Form Controllable Text Generation

DGX agent

arXiv:2508.04047v2 Announce Type: replace Abstract: Prefix-based methods have emerged as a promising paradigm for Controllable Text Generation (CTG) due to their parameter efficiency. However, while e

model-releasesarxiv-cs-cl
19 May 2026
Research

Large Language Models and Impossible Language Acquisition: 'False Promise' or an Overturn of our Current Perspective towards AI

DGX agent

arXiv:2602.08437v5 Announce Type: replace Abstract: In Chomsky's provocative critique 'The False Promise of CHATGPT,' Large Language Models (LLMs) are characterized as mere pattern predictors that do

researcharxiv-cs-cl
19 May 2026
Agents

Learning from Self-Debate: Preparing Reasoning Models for Multi-Agent Debate

DGX agent

arXiv:2601.22297v2 Announce Type: replace Abstract: The reasoning abilities of large language models (LLMs) have been substantially improved by reinforcement learning with verifiable rewards (RLVR). A

agentsarxiv-cs-cl
19 May 2026
Safety

Learning to Reason without External Rewards

DGX agent

arXiv:2505.19590v5 Announce Type: replace-cross Abstract: Training large language models (LLMs) for complex reasoning via Reinforcement Learning with Verifiable Rewards (RLVR) is effective but limited

safetyarxiv-cs-cl
19 May 2026
Safety

Learning Transferable Topology Priors for Multi-Agent LLM Collaboration Across Domains

DGX agent

arXiv:2605.17359v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems have shown strong potential for complex reasoning by coordinating specialized agents through struct

safetyarxiv-cs-cl
19 May 2026
Research

Leveraging Multimodal Self-Consistency Reasoning in Coding Motivational Interviewing for Alcohol Use Reduction

DGX agent

arXiv:2605.12987v2 Announce Type: replace Abstract: BACKGROUND: Coding Motivational Interviewing (MI) sessions is essential for understanding client behaviors and predicting outcomes, but it requires

researcharxiv-cs-cl
19 May 2026
Safety

Linguistic Uncertainty and Reply Engagement on X: A Cross-Domain Replication of the Uncertainty-Reply Asymmetry

DGX agent

arXiv:2605.16289v1 Announce Type: cross Abstract: Linguistic uncertainty is common in social media, but its relationship with engagement remains unclear across languages and topics. Using 2,258 Englis

safetyarxiv-cs-cl
19 May 2026
Safety

LISTEN to Your Preferences: An LLM Framework for Multi-Objective Selection

DGX agent

arXiv:2510.25799v2 Announce Type: replace Abstract: Human experts often struggle to select the best option from a large set of items with multiple competing objectives, a process bottlenecked by the d

safetyarxiv-cs-cl
19 May 2026
Applications

LLM Agents Are the Antidote to Walled Gardens

DGX agent

arXiv:2506.23978v3 Announce Type: replace-cross Abstract: While the Internet's core infrastructure was designed to be open and universal, today's application layer is dominated by closed, proprietary

applicationsarxiv-cs-cl
19 May 2026
Tutorials

LLM-Based Intelligent Notification Composition: From Static Personalization to Context-Aware Persuasive Messaging

DGX agent

arXiv:2605.16264v1 Announce Type: cross Abstract: Push notifications remain among the most direct channels through which digital platforms engage users, yet existing approaches have invested heavily i

tutorialsarxiv-cs-cl
19 May 2026
Local Ai

LLMs for automatic annotation of Mandarin narrative transcripts

DGX agent

arXiv:2605.17205v1 Announce Type: new Abstract: Linguistic annotation of transcribed speech is essential for research in language acquisition, language disorders, and sociolinguistics, yet remains lab

local-aiarxiv-cs-cl
19 May 2026
Model Releases

LLMs in Qualitative Research: Opportunities, Limitations, and Practical Considerations

DGX agent

arXiv:2605.16538v1 Announce Type: cross Abstract: This paper examines the opportunities, limitations, and practical considerations associated with the use of large language models (LLMs) in qualitativ

model-releasesarxiv-cs-cl
19 May 2026
Agents

MA^{2}P: A Meta-Cognitive Autonomous Intelligent Agents Framework for Complex Persuasion

DGX agent

arXiv:2605.18572v1 Announce Type: new Abstract: Persuasive dialogue generation plays a vital role in decision-making, negotiation, counseling, and behavior change, yet it remains a challenging problem

agentsarxiv-cs-cl
19 May 2026
Safety

Medical Context Distorts Decisions in Clinical Vision Language Models

DGX agent

arXiv:2605.17436v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly proposed for clinical decision support, yet their reliability in real-world scenarios that require inte

safetyarxiv-cs-cl
19 May 2026
Model Releases

MentalBench: A DSM-Grounded Benchmark for Evaluating Psychiatric Diagnostic Capability of Large Language Models

DGX agent

arXiv:2602.12871v2 Announce Type: replace Abstract: Large language models (LLMs) have attracted growing interest as supportive tools for psychiatric assessment and clinical decision support. However,

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Merlin's Whisper: Enabling Efficient Reasoning in Large Language Models via Black-box Persuasive Prompting

DGX agent

arXiv:2510.10528v3 Announce Type: replace Abstract: Large reasoning models (LRMs) have demonstrated remarkable proficiency in tackling complex tasks through step-by-step thinking. However, this length

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

MiniGPT: Rebuilding GPT from First Principles

DGX agent

arXiv:2605.17398v1 Announce Type: new Abstract: This paper presents MiniGPT, a compact from-scratch implementation of GPT-style autoregressive language modeling in PyTorch. The aim is to rebuild the c

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

MixSD: Mixed Contextual Self-Distillation for Knowledge Injection

DGX agent

arXiv:2605.16865v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is widely used to inject new knowledge into language models, but it often degrades pretrained capabilities such as reasonin

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Mixture of Experts for Low-Resource LLMs

DGX agent

arXiv:2605.17598v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures enable efficient model scaling, yet expert routing behavior across underrepresented languages remains poorly unde

model-releasesarxiv-cs-cl
19 May 2026
← Previous
1…8990919293…162
Next →