AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

LoRA-FA: Efficient and Effective Low Rank Representation Fine-tuning

DGX agent

arXiv:2308.03303v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) is crucial for improving their performance on downstream tasks, but full-parameter fine-tuning (Full-FT) is

model-releasesarxiv-cs-cl
23 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Markov reads Pushkin, again: A statistical journey into the poetic world of Evgenij Onegin

DGX agent

arXiv:2604.20221v1 Announce Type: new Abstract: This study applies symbolic time series analysis and Markov modeling to explore the phonological structure of Evgenij Onegin-as captured through a graph

researcharxiv-cs-cl
23 Apr 2026
Research

Mechanistic Interpretability of Large-Scale Counting in LLMs through a System-2 Strategy

DGX agent

arXiv:2601.02989v2 Announce Type: replace Abstract: Large language models (LLMs), despite strong performance on complex mathematical problems, exhibit systematic limitations in counting tasks. This is

researcharxiv-cs-cl
23 Apr 2026
Safety

Memorization, Emergence, and Explaining Reversal Failures: A Controlled Study of Relational Semantics in LLMs

DGX agent

arXiv:2601.02931v2 Announce Type: replace Abstract: Autoregressive LLMs perform well on relational tasks that require linking entities via relational words (e.g., father/son, friend), but it is unclea

safetyarxiv-cs-cl
23 Apr 2026
Safety

MOA: Multi-Objective Alignment for Role-Playing Agents

DGX agent

arXiv:2512.09756v2 Announce Type: replace Abstract: Role-playing agents (RPAs) require balancing multiple objectives, such as instruction following, persona consistency, and stylistic fidelity, which

safetyarxiv-cs-cl
23 Apr 2026
Research

Model Internal Sleuthing: Finding Lexical Identity and Inflectional Features in Modern Language Models

DGX agent

arXiv:2506.02132v5 Announce Type: replace Abstract: Large transformer-based language models dominate modern NLP, yet our understanding of how they encode linguistic information relies primarily on stu

researcharxiv-cs-cl
23 Apr 2026
Research

Multi-Perspective Evidence Synthesis and Reasoning for Unsupervised Multimodal Entity Linking

DGX agent

arXiv:2604.20283v1 Announce Type: new Abstract: Multimodal Entity Linking (MEL) is a fundamental task in data management that maps ambiguous mentions with diverse modalities to the multimodal entities

researcharxiv-cs-cl
23 Apr 2026
Agents

Neural Bandit Based Optimal LLM Selection for a Pipeline of Subtasks

DGX agent

arXiv:2508.09958v3 Announce Type: replace Abstract: As large language models (LLMs) become increasingly popular, there is a growing need to predict which out of a set of LLMs will yield a successful a

agentsarxiv-cs-cl
23 Apr 2026
Model Releases

'Newspaper Eat' Means 'Not Tasty': A Taxonomy and Benchmark for Coded Language in Real-World Chinese Online Reviews

DGX agent

arXiv:2601.19932v2 Announce Type: replace Abstract: Coded language is an important part of human communication. It refers to cases where users intentionally encode meaning so that the surface text dif

model-releasesarxiv-cs-cl
23 Apr 2026
Research

Not all ANIMALs are equal: metaphorical framing through source domains and semantic frames

DGX agent

arXiv:2604.20454v1 Announce Type: new Abstract: Metaphors are powerful framing devices, yet their source domains alone do not fully explain the specific associations they evoke. We argue that the inte

researcharxiv-cs-cl
23 Apr 2026
Research

On the Quantization Robustness of Diffusion Language Models in Coding Benchmarks

DGX agent

arXiv:2604.20079v1 Announce Type: cross Abstract: Auto-regressive Large Language Models (LLMs) achieve strong performance on coding tasks, but incur high memory and inference costs. Diffusion-based la

researcharxiv-cs-cl
23 Apr 2026
Research

Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization

DGX agent

arXiv:2601.12078v2 Announce Type: replace Abstract: Large language models (LLMs) excel at general-purpose tasks, yet adapting their responses to individual users remains challenging. Retrieval augment

researcharxiv-cs-cl
23 Apr 2026
Research

Over-Refusal and Representation Subspaces: A Mechanistic Analysis of Task-Conditioned Refusal in Aligned LLMs

DGX agent

arXiv:2603.27518v2 Announce Type: replace Abstract: Aligned language models that are trained to refuse harmful requests also exhibit over-refusal: they decline safe instructions that seemingly resembl

researcharxiv-cs-cl
23 Apr 2026
Model Releases

Parallel-SFT: Improving Zero-Shot Cross-Programming-Language Transfer for Code RL

DGX agent

arXiv:2604.20835v1 Announce Type: new Abstract: Modern language models demonstrate impressive coding capabilities in common programming languages (PLs), such as C++ and Python, but their performance i

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

PLR: Plackett-Luce for Reordering In-Context Learning Examples

DGX agent

arXiv:2603.21373v2 Announce Type: replace-cross Abstract: In-context learning (ICL) adapts large language models by conditioning on a small set of ICL examples, avoiding costly parameter updates. Amon

model-releasesarxiv-cs-cl
23 Apr 2026
Applications

RADS: Reinforcement Learning-Based Sample Selection Improves Transfer Learning in Low-resource and Imbalanced Clinical Settings

DGX agent

arXiv:2604.20256v1 Announce Type: new Abstract: A common strategy in transfer learning is few shot fine-tuning, but its success is highly dependent on the quality of samples selected as training examp

applicationsarxiv-cs-cl
23 Apr 2026
Model Releases

RespondeoQA: a Benchmark for Bilingual Latin-English Question Answering

DGX agent

arXiv:2604.20738v1 Announce Type: new Abstract: We introduce a benchmark dataset for question answering and translation in bilingual Latin and English settings, containing about 7,800 question-answer

model-releasesarxiv-cs-cl
23 Apr 2026
Safety

Rethinking Reinforcement Fine-Tuning in LVLM: Convergence, Reward Decomposition, and Generalization

DGX agent

arXiv:2604.19857v1 Announce Type: cross Abstract: Reinforcement fine-tuning with verifiable rewards (RLVR) has emerged as a powerful paradigm for equipping large vision-language models (LVLMs) with ag

safetyarxiv-cs-cl
23 Apr 2026
Research

Retrofitting Small Multilingual Models for Retrieval: Matching 7B Performance with 300M Parameters

DGX agent

arXiv:2510.14274v2 Announce Type: replace Abstract: Training effective multilingual embedding models presents unique challenges due to the diversity of languages and task objectives. Although small mu

researcharxiv-cs-cl
23 Apr 2026
Model Releases

RExBench: Can coding agents autonomously implement AI research extensions?

DGX agent

arXiv:2506.22598v3 Announce Type: replace Abstract: Agents based on Large Language Models (LLMs) have shown promise for performing sophisticated software engineering tasks autonomously. In addition, t

model-releasesarxiv-cs-cl
23 Apr 2026
Local Ai

SAKE: Self-aware Knowledge Exploitation-Exploration for Grounded Multimodal Named Entity Recognition

DGX agent

arXiv:2604.20146v1 Announce Type: cross Abstract: Grounded Multimodal Named Entity Recognition (GMNER) aims to extract named entities and localize their visual regions within image-text pairs, serving

local-aiarxiv-cs-cl
23 Apr 2026
Model Releases

Self-Aware Vector Embeddings for Retrieval-Augmented Generation: A Neuroscience-Inspired Framework for Temporal, Confidence-Weighted, and Relational Knowledge

DGX agent

arXiv:2604.20598v1 Announce Type: cross Abstract: Modern retrieval-augmented generation (RAG) systems treat vector embeddings as static, context-free artifacts: an embedding has no notion of when it w

model-releasesarxiv-cs-cl
23 Apr 2026
Safety

SignDATA: Data Pipeline for Sign Language Translation

DGX agent

arXiv:2604.20357v1 Announce Type: cross Abstract: Sign-language datasets are difficult to preprocess consistently because they vary in annotation schema, clip timing, signer framing, and privacy const

safetyarxiv-cs-cl
23 Apr 2026
Model Releases

SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks

DGX agent

arXiv:2604.20087v1 Announce Type: new Abstract: Skills have become the de facto way to enable LLM agents to perform complex real-world tasks with customized instructions, workflows, and tools, but how

model-releasesarxiv-cs-cl
23 Apr 2026
Applications

Structured Disagreement in Health-Literacy Annotation: Epistemic Stability, Conceptual Difficulty, and Agreement-Stratified Inference

DGX agent

arXiv:2604.19943v1 Announce Type: new Abstract: Annotation pipelines in Natural Language Processing (NLP) commonly assume a single latent ground truth per instance and resolve disagreement through lab

applicationsarxiv-cs-cl
23 Apr 2026
Research

Task-Dependent Evaluation of LLM Output Homogenization: A Taxonomy-Guided Framework

DGX agent

arXiv:2509.21267v3 Announce Type: replace Abstract: Large language models often generate homogeneous outputs, but whether this is problematic depends on the specific task. For objective math tasks, re

researcharxiv-cs-cl
23 Apr 2026
Model Releases

Text-to-Distribution Prediction with Quantile Tokens and Neighbor Context

DGX agent

arXiv:2604.20216v1 Announce Type: new Abstract: Many applications of LLM-based text regression require predicting a full conditional distribution rather than a single point value. We study distributio

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

The GaoYao Benchmark: A Comprehensive Framework for Evaluating Multilingual and Multicultural Abilities of Large Language Models

DGX agent

arXiv:2604.20225v1 Announce Type: new Abstract: Evaluating the multilingual and multicultural capabilities of Large Language Models (LLMs) is essential for their global utility. However, current bench

model-releasesarxiv-cs-cl
23 Apr 2026
Safety

The Imperfective Paradox in Large Language Models

DGX agent

arXiv:2601.09373v2 Announce Type: replace Abstract: Do Large Language Models (LLMs) genuinely grasp the compositional semantics of events, or do they rely on surface-level probabilistic heuristics? We

safetyarxiv-cs-cl
23 Apr 2026
Model Releases

To Know is to Construct: Schema-Constrained Generation for Agent Memory

DGX agent

arXiv:2604.20117v1 Announce Type: new Abstract: Constructivist epistemology argues that knowledge is actively constructed rather than passively copied. Despite the generative nature of Large Language

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Towards High-Quality Machine Translation for Kokborok: A Low-Resource Tibeto-Burman Language of Northeast India

DGX agent

arXiv:2604.19778v1 Announce Type: new Abstract: We present KokborokMT, a high-quality neural machine translation (NMT) system for Kokborok (ISO 639-3), a Tibeto-Burman language spoken primarily in Tri

model-releasesarxiv-cs-cl
23 Apr 2026
Research

Tracing Relational Knowledge Recall in Large Language Models

DGX agent

arXiv:2604.19934v1 Announce Type: new Abstract: We study how large language models recall relational knowledge during text generation, with a focus on identifying latent representations suitable for r

researcharxiv-cs-cl
23 Apr 2026
Model Releases

Trajectory2Task: Training Robust Tool-Calling Agents with Synthesized Yet Verifiable Data for Complex User Intents

DGX agent

arXiv:2601.20144v3 Announce Type: replace Abstract: Tool-calling agents are increasingly deployed in real-world customer-facing workflows. Yet most studies on tool-calling agents focus on idealized se

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

WebGen-R1: Incentivizing Large Language Models to Generate Functional and Aesthetic Websites with Reinforcement Learning

DGX agent

arXiv:2604.20398v1 Announce Type: new Abstract: While Large Language Models (LLMs) excel at function-level code generation, project-level tasks such as generating functional and visually aesthetic mul

model-releasesarxiv-cs-cl
23 Apr 2026
Applications

What Language Models Know But Don't Say: Non-Generative Prior Extraction for Generalization

DGX agent

arXiv:2601.17609v2 Announce Type: replace Abstract: In domains like medicine and finance, large-scale labeled data is costly and often unavailable, leading to models trained on small datasets that str

applicationsarxiv-cs-cl
23 Apr 2026
Local Ai

Where Reasoning Breaks: Logic-Aware Path Selection by Controlling Logical Connectives in LLMs Reasoning Chains

DGX agent

arXiv:2604.20564v1 Announce Type: new Abstract: While LLMs demonstrate impressive reasoning capabilities, they remain fragile in multi-step logical deduction, where a single transition error can propa

local-aiarxiv-cs-cl
23 Apr 2026
Safety

Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Alignment

DGX agent

arXiv:2601.14249v4 Announce Type: replace Abstract: Long chain-of-thought (CoT) trajectories provide rich supervision signals for distilling reasoning from teacher to student LLMs. However, both prior

safetyarxiv-cs-cl
23 Apr 2026
Safety

Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives

DGX agent

arXiv:2604.20131v1 Announce Type: new Abstract: Increasingly, studies are exploring using Large Language Models (LLMs) for accelerated or scaled qualitative analysis of text data. While we can compare

safetyarxiv-cs-cl
23 Apr 2026
Research

WISCA: A Lightweight Model Transition Method to Improve LLM Training via Weight Scaling

DGX agent

arXiv:2508.16676v2 Announce Type: replace-cross Abstract: Transformer architecture gradually dominates the LLM field. Recent advances in training optimization for Transformer-based large language mode

researcharxiv-cs-cl
23 Apr 2026
Applications

A Bolu: A Structured Dataset for the Computational Analysis of Sardinian Improvisational Poetry

DGX agent

arXiv:2604.19584v1 Announce Type: new Abstract: The growing interest of Natural Language Processing (NLP) in minority languages has not yet bridged the gap in the preservation of oral linguistic herit

applicationsarxiv-cs-cl
22 Apr 2026
Research

A Mechanism and Optimization Study on the Impact of Information Density on User-Generated Content Named Entity Recognition

DGX agent

arXiv:2604.18944v1 Announce Type: new Abstract: Named Entity Recognition (NER) models trained on clean, high-resource corpora exhibit catastrophic performance collapse when deployed on noisy, sparse U

researcharxiv-cs-cl
22 Apr 2026
Agents

A Self-Evolving Framework for Efficient Terminal Agents via Observational Context Compression

DGX agent

arXiv:2604.19572v1 Announce Type: new Abstract: As model capabilities advance, research has increasingly shifted toward long-horizon, multi-turn terminal-centric agentic tasks, where raw environment f

agentsarxiv-cs-cl
22 Apr 2026
Model Releases

AlignCultura: Towards Culturally Aligned Large Language Models?

DGX agent

arXiv:2604.19016v1 Announce Type: new Abstract: Cultural alignment in Large Language Models (LLMs) is essential for producing contextually aware, respectful, and trustworthy outputs. Without it, model

model-releasesarxiv-cs-cl
22 Apr 2026
Research

An Answer is just the Start: Related Insight Generation for Open-Ended Document-Grounded QA

DGX agent

arXiv:2604.19685v1 Announce Type: new Abstract: Answering open-ended questions remains challenging for AI systems because it requires synthesis, judgment, and exploration beyond factual retrieval, and

researcharxiv-cs-cl
22 Apr 2026
Model Releases

An Empirical Study of Multi-Generation Sampling for Jailbreak Detection in Large Language Models

DGX agent

arXiv:2604.18775v1 Announce Type: new Abstract: Detecting jailbreak behaviour in large language models remains challenging, particularly when strongly aligned models produce harmful outputs only rarel

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Are Large Language Models Economically Viable for Industry Deployment?

DGX agent

arXiv:2604.19342v1 Announce Type: new Abstract: Generative AI-powered by Large Language Models (LLMs)-is increasingly deployed in industry across healthcare decision support, financial analytics, ente

model-releasesarxiv-cs-cl
22 Apr 2026
Research

Article and Comment Frames Shape the Quality of Online Comments

DGX agent

arXiv:2603.27889v2 Announce Type: replace Abstract: Framing theory posits that how information is presented shapes audience responses, but computational work has largely ignored audience reactions. Wh

researcharxiv-cs-cl
22 Apr 2026
Model Releases

Bangla Key2Text: Text Generation from Keywords for a Low Resource Language

DGX agent

arXiv:2604.19508v1 Announce Type: new Abstract: This paper introduces extit{Bangla Key2Text}, a large-scale dataset of 2.6 million Bangla keyword--text pairs designed for keyword-driven text generatio

model-releasesarxiv-cs-cl
22 Apr 2026
← Previous
1…127128129130131…161
Next →