AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

Copy First, Translate Later: Interpreting Translation Dynamics in Multilingual Pretraining

DGX agent

arXiv:2604.17633v1 Announce Type: new Abstract: Large language models exhibit impressive cross-lingual capabilities. However, prior work analyzes this phenomenon through isolated factors and at sparse

model-releasesarxiv-cs-cl
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

COSEARCH: Joint Training of Reasoning and Document Ranking via Reinforcement Learning for Agentic Search

DGX agent

arXiv:2604.17555v1 Announce Type: cross Abstract: Agentic search -- the task of training agents that iteratively reason, issue queries, and synthesize retrieved information to answer complex questions

safetyarxiv-cs-cl
21 Apr 2026
Research

Countdown-Code: A Testbed for Studying The Emergence and Generalization of Reward Hacking in RLVR

DGX agent

arXiv:2603.07084v2 Announce Type: replace-cross Abstract: Reward hacking is a form of misalignment in which models overoptimize proxy rewards without genuinely solving the underlying task. Precisely m

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Creating ConLangs to Probe the Metalinguistic Grammatical Knowledge of LLMs

DGX agent

arXiv:2510.07591v3 Announce Type: replace Abstract: We present a system that uses LLMs as a tool in the development of Constructed Languages -- ConLangs, which we call IASC (Interactive Agentic System

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CreditDecoding: Accelerating Parallel Decoding in Diffusion Large Language Models with Trace Credit

DGX agent

arXiv:2510.06133v2 Announce Type: replace Abstract: Diffusion large language models (dLLMs) generate text through iterative denoising. In commonly adopted parallel decoding schemes, each step confirms

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

CRISP: Compressing Redundancy in Chain-of-Thought via Intrinsic Saliency Pruning

DGX agent

arXiv:2604.17297v1 Announce Type: new Abstract: Long Chain-of-Thought (CoT) reasoning is pivotal for the success of recent reasoning models but suffers from high computational overhead and latency. Wh

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

CROC: Evaluating and Training T2I Metrics with Pseudo- and Human-Labeled Contrastive Robustness Checks

DGX agent

arXiv:2505.11314v2 Announce Type: replace-cross Abstract: The assessment of evaluation metrics (meta-evaluation) is crucial for determining the suitability of existing metrics in text-to-image (T2I) g

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Cross-Family Speculative Decoding for Polish Language Models on Apple~Silicon: An Empirical Evaluation of Bielik~11B with UAG-Extended MLX-LM

DGX agent

arXiv:2604.16368v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by using a small draft model to propose k candidate tokens for a target model to verify. While effective

model-releasesarxiv-cs-cl
21 Apr 2026
Applications

Crowded in B-Space: Calibrating Shared Directions for LoRA Merging

DGX agent

arXiv:2604.16826v1 Announce Type: new Abstract: Merging separately trained LoRA adapters is a practical alternative to joint multi-task training, but it often hurts performance. Existing methods usual

applicationsarxiv-cs-cl
21 Apr 2026
Applications

CT Open: An Open-Access, Uncontaminated, Live Platform for the Open Challenge of Clinical Trial Outcome Prediction

DGX agent

arXiv:2604.16742v1 Announce Type: cross Abstract: Scientists have long sought to accurately predict outcomes of real-world events before they happen. Can AI systems do so more reliably? We study this

applicationsarxiv-cs-cl
21 Apr 2026
Research

Culinary Crossroads: A RAG Framework for Enhancing Diversity in Cross-Cultural Recipe Adaptation

DGX agent

arXiv:2507.21934v2 Announce Type: replace Abstract: In cross-cultural recipe adaptation, the goal is not only to ensure cultural appropriateness and retain the original dish's essence, but also to pro

researcharxiv-cs-cl
21 Apr 2026
Safety

Culture-Aware Humorous Captioning: Multimodal Humor Generation across Cultural Contexts

DGX agent

arXiv:2604.18091v1 Announce Type: new Abstract: Recent multimodal large language models have shown promising ability in generating humorous captions for images, yet they still lack stable control over

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

DART: Mitigating Harm Drift in Difference-Aware LLMs via Distill-Audit-Repair Training

DGX agent

arXiv:2604.16845v1 Announce Type: new Abstract: Large language models (LLMs) tuned for safety often avoid acknowledging demographic differences, even when such acknowledgment is factually correct (e.g

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Data Compressibility Quantifies LLM Memorization

DGX agent

arXiv:2507.06056v4 Announce Type: replace Abstract: Large Language Models (LLMs) are known to memorize portions of their training data, sometimes even reproduce content verbatim when prompted appropri

researcharxiv-cs-cl
21 Apr 2026
Research

Data Mixing for Large Language Models Pretraining: A Survey and Outlook

DGX agent

arXiv:2604.16380v1 Announce Type: new Abstract: Large language models (LLMs) rely on pretraining on massive and heterogeneous corpora, where training data composition has a decisive impact on training

researcharxiv-cs-cl
21 Apr 2026
Research

Decisive: Guiding User Decisions with Optimal Preference Elicitation from Unstructured Documents

DGX agent

arXiv:2604.18122v1 Announce Type: new Abstract: Decision-making is a cognitively intensive task that requires synthesizing relevant information from multiple unstructured sources, weighing competing f

researcharxiv-cs-cl
21 Apr 2026
Safety

Decoupling the Effect of Chain-of-Thought Reasoning: A Human Label Variation Perspective

DGX agent

arXiv:2601.03154v2 Announce Type: replace Abstract: Reasoning-tuned LLMs utilizing long Chain-of-Thought (CoT) excel at single-answer tasks, yet their ability to model Human Label Variation--which req

safetyarxiv-cs-cl
21 Apr 2026
Research

Defragmenting Language Models: An Interpretability-based Approach for Vocabulary Expansion

DGX agent

arXiv:2604.16656v1 Announce Type: new Abstract: All languages are equal; when it comes to tokenization, some are more equal than others. Tokens are the hidden currency that dictate the cost and latenc

researcharxiv-cs-cl
21 Apr 2026
Research

DeInfer: Efficient Parallel Inferencing for Decomposed Large Language Models

DGX agent

arXiv:2604.17709v1 Announce Type: new Abstract: Existing works on large language model (LLM) decomposition mainly focus on improving performance on downstream tasks, but they ignore the poor parallel

researcharxiv-cs-cl
21 Apr 2026
Safety

Demystifying the unreasonable effectiveness of online alignment methods

DGX agent

arXiv:2604.17207v1 Announce Type: cross Abstract: Iterative alignment methods based on purely greedy updates are remarkably effective in practice, yet existing theoretical guarantees of (O(log T)) KL-

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Depth Registers Unlock W4A4 on SwiGLU: A Reader/Generator Decomposition

DGX agent

arXiv:2604.18128v1 Announce Type: new Abstract: We study post-training W4A4 quantization in a controlled 300M-parameter SwiGLU decoder-only language model trained on 5B tokens of FineWeb-Edu, and ask

model-releasesarxiv-cs-cl
21 Apr 2026
Agents

Designing Explainable Conversational Agentic Systems for Guarani Speakers

DGX agent

arXiv:2603.05743v3 Announce Type: replace Abstract: Although artificial intelligence (AI) and Human-Computer Interaction (HCI) systems are often presented as universal solutions, their design remains

agentsarxiv-cs-cl
21 Apr 2026
Safety

Detecting Alarming Student Verbal Responses using Text and Audio Classifier

DGX agent

arXiv:2604.16717v1 Announce Type: new Abstract: This paper addresses a critical safety gap in the use Automated Verbal Response Scoring (AVRS). We present a novel hybrid framework for troubled student

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Detecting LLM-Generated Spam Reviews by Integrating Language Model Embeddings and Graph Neural Network

DGX agent

arXiv:2510.01801v2 Announce Type: replace Abstract: The rise of large language models (LLMs) has enabled the generation of highly persuasive spam reviews that closely mimic human writing. These review

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Diagnosing LLM-based Rerankers in Cold-Start Recommender Systems: Coverage, Exposure and Practical Mitigations

DGX agent

arXiv:2604.16318v1 Announce Type: cross Abstract: Large language models (LLMs) and cross-encoder rerankers have gained attention for improving recommender systems, particularly in cold-start scenarios

safetyarxiv-cs-cl
21 Apr 2026
Safety

DiffCoT: Diffusion-styled Chain-of-Thought Reasoning in LLMs

DGX agent

arXiv:2601.03559v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning improves multi-step mathematical problem solving in large language models but remains vulnerable to exposure bias a

safetyarxiv-cs-cl
21 Apr 2026
Local Ai

Different Paths to Harmful Compliance: Behavioral Side Effects and Mechanistic Divergence Across LLM Jailbreaks

DGX agent

arXiv:2604.18510v1 Announce Type: cross Abstract: Open-weight language models can be rendered unsafe through several distinct interventions, but the resulting models may differ substantially in capabi

local-aiarxiv-cs-cl
21 Apr 2026
Agents

Diversity Collapse in Multi-Agent LLM Systems: Structural Coupling and Collective Failure in Open-Ended Idea Generation

DGX agent

arXiv:2604.18005v1 Announce Type: cross Abstract: Multi-agent systems (MAS) are increasingly used for open-ended idea generation, driven by the expectation that collective interaction will broaden the

agentsarxiv-cs-cl
21 Apr 2026
Research

Do LLMs Encode Functional Importance of Reasoning Tokens?

DGX agent

arXiv:2601.03066v2 Announce Type: replace Abstract: Large language models solve complex tasks by generating long reasoning chains, achieving higher accuracy at the cost of increased computational cost

researcharxiv-cs-cl
21 Apr 2026
Safety

Do LLMs Use Cultural Knowledge Without Being Told? A Multilingual Evaluation of Implicit Pragmatic Adaptation

DGX agent

arXiv:2604.17718v1 Announce Type: new Abstract: Many benchmarks show that large language models can answer direct questions about culture. We study a different question: do they also change how they s

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

DocQAC: Adaptive Trie-Guided Decoding for Effective In-Document Query Auto-Completion

DGX agent

arXiv:2604.18257v1 Announce Type: cross Abstract: Query auto-completion (QAC) has been widely studied in the context of web search, yet remains underexplored for in-document search, which we term DocQ

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Document-as-Image Representations Fall Short for Scientific Retrieval

DGX agent

arXiv:2604.18508v1 Announce Type: cross Abstract: Many recent document embedding models are trained on document-as-image representations, embedding rendered pages as images rather than the underlying

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Does Welsh media need a review? Detecting bias in Nation.Cymru's political reporting

DGX agent

arXiv:2604.17628v1 Announce Type: new Abstract: Wales' political landscape has been marked by growing accusations of bias in Welsh media. This paper takes the first computational step toward testing t

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Domain-oriented RAG Assessment (DoRA): Synthetic Benchmarking for RAG-based Question Answering on Defense Documents

DGX agent

arXiv:2604.17943v1 Announce Type: new Abstract: Open-domain RAG benchmarks over public corpora can overestimate deployment performance due to pretraining overlap and weak attribution requirements. We

model-releasesarxiv-cs-cl
21 Apr 2026
Agents

Don't Adapt Small Language Models for Tools; Adapt Tool Schemas to the Models

DGX agent

arXiv:2510.07248v3 Announce Type: replace Abstract: Small language models (SLMs) enable scalable tool-augmented multi-agent systems where multiple SLMs handle subtasks orchestrated by a powerful coord

agentsarxiv-cs-cl
21 Apr 2026
Model Releases

DORA Explorer: Improving the Exploration Ability of LLMs Without Training

DGX agent

arXiv:2604.17244v1 Announce Type: new Abstract: Despite the rapid progress, LLMs for sequential decision-making (i.e., LLM agents) still struggle to produce diverse outputs. This leads to insufficient

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

DOSE: Data Selection for Multi-Modal LLMs via Off-the-Shelf Models

DGX agent

arXiv:2604.16979v1 Announce Type: cross Abstract: High-quality and diverse multimodal data are essential for improving vision-language models (VLMs), yet existing datasets often contain noisy, redunda

safetyarxiv-cs-cl
21 Apr 2026
Safety

Dual Alignment Between Language Model Layers and Human Sentence Processing

DGX agent

arXiv:2604.18563v1 Announce Type: new Abstract: A recent study (Kuribayashi et al., 2025) has shown that human sentence processing behavior, typically measured on syntactically unchallenging construct

safetyarxiv-cs-cl
21 Apr 2026
Research

DualToken: Towards Unifying Visual Understanding and Generation with Dual Visual Vocabularies

DGX agent

arXiv:2503.14324v3 Announce Type: replace-cross Abstract: The differing representation spaces required for visual understanding and generation pose a challenge in unifying them within the autoregressi

researcharxiv-cs-cl
21 Apr 2026
Model Releases

DuConTE: Dual-Granularity Text Encoder with Topology-Constrained Attention for Text-attributed Graphs

DGX agent

arXiv:2604.17411v1 Announce Type: new Abstract: Text-attributed graphs integrate semantic information of node texts with topological structure, offering significant value in various applications such

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

DuQuant++: Fine-grained Rotation Enhances Microscaling FP4 Quantization

DGX agent

arXiv:2604.17789v1 Announce Type: cross Abstract: The MXFP4 microscaling format, which partitions tensors into blocks of 32 elements sharing an E8M0 scaling factor, has emerged as a promising substrat

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Dynamic Emotion and Personality Profiling for Multimodal Deception Detection

DGX agent

arXiv:2604.17037v1 Announce Type: new Abstract: Deception detection is of great significance for ensuring information security and conducting public opinion analysis, with personality factors and emot

safetyarxiv-cs-cl
21 Apr 2026
Safety

DynaWeb: Model-Based Reinforcement Learning of Web Agents

DGX agent

arXiv:2601.22149v2 Announce Type: replace Abstract: The development of autonomous web agents, powered by Large Language Models (LLMs) and reinforcement learning (RL), represents a significant step tow

safetyarxiv-cs-cl
21 Apr 2026
Research

E2E-GMNER: End-to-End Generative Grounded Multimodal Named Entity Recognition

DGX agent

arXiv:2604.17319v1 Announce Type: cross Abstract: Grounded Multimodal Named Entity Recognition (GMNER) aims to jointly identify named entity mentions in text, predict their semantic types, and ground

researcharxiv-cs-cl
21 Apr 2026
Model Releases

EchoChain: A Full-Duplex Benchmark for State-Update Reasoning Under Interruptions

DGX agent

arXiv:2604.16456v1 Announce Type: new Abstract: Real-time voice assistants must revise task state when users interrupt mid-response, but existing spoken-dialog benchmarks largely evaluate turn-based i

model-releasesarxiv-cs-cl
21 Apr 2026
Tutorials

EduRABSA: An Education Review Dataset for Aspect-based Sentiment Analysis Tasks

DGX agent

arXiv:2508.17008v2 Announce Type: replace Abstract: Every year, most educational institutions seek and receive an enormous volume of text feedback from students on courses, teaching, and overall exper

tutorialsarxiv-cs-cl
21 Apr 2026
Research

Efficient Low-Resource Language Adaptation via Multi-Source Dynamic Logit Fusion

DGX agent

arXiv:2604.18106v1 Announce Type: new Abstract: Adapting large language models (LLMs) to low-resource languages (LRLs) is constrained by the scarcity of task data and computational resources. Although

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Efficient Task Adaptation in Large Language Models via Selective Parameter Optimization

DGX agent

arXiv:2604.17051v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated excellent performance in general language understanding, generation and other tasks. However, when fine-t

model-releasesarxiv-cs-cl
21 Apr 2026
← Previous
1…132133134135136…161
Next →