AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

LMs as Task-Specific Knowledge Bases: An Interpretability Analysis

DGX agent

arXiv:2606.27237v1 Announce Type: new Abstract: Language models (LMs) capture large amounts of factual knowledge applicable to a wide range of tasks, motivating the view of their parameters as a knowl

model-releasesarxiv-cs-cl
26 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Applications

Mapping Political-Elite Networks in Europe with a Multilingual Joint Entity-Relation Extraction Pipeline

DGX agent

arXiv:2606.27347v1 Announce Type: new Abstract: Whether political elites organise into rent-seeking coalitions that capture public resources or civic networks that sustain governance is a central ques

applicationsarxiv-cs-cl
26 Jun 2026
Safety

MinGram: A Minimalist Unigram Tokenizer with High Compression and Competitive Morphological Alignment

DGX agent

arXiv:2606.27019v1 Announce Type: new Abstract: The Unigram tokenizer uses an elegant representation which makes it straightforward to edit vocabularies, but its training is comparatively heavy and co

safetyarxiv-cs-cl
26 Jun 2026
Research

Multilingual Reasoning Cascades Need More Context

DGX agent

arXiv:2606.27306v1 Announce Type: new Abstract: Translation cascades for reasoning translate the query from another language to English, reason in English, and translate the answer back to the origina

researcharxiv-cs-cl
26 Jun 2026
Model Releases

Nemotron-TwoTower: Diffusion Language Modeling with Pretrained Autoregressive Context

DGX agent

arXiv:2606.26493v1 Announce Type: new Abstract: Diffusion language models offer a promising alternative to autoregressive models due to their potential for parallel and iterative generation. However,

model-releasesarxiv-cs-cl
26 Jun 2026
Safety

Neural Speaker Diarization via Multilingual Training: Evaluation on Low-Resource Nepali-Hindi Speech

DGX agent

arXiv:2606.26144v1 Announce Type: cross Abstract: Speaker diarization, the task of determining 'who spoke when' in a multi-speaker recording, is a critical component in applications such as meeting tr

safetyarxiv-cs-cl
26 Jun 2026
Model Releases

OI-Bench: An Option Injection Benchmark for Evaluating LLM Susceptibility to Directive Interference

DGX agent

arXiv:2601.13300v2 Announce Type: replace Abstract: Benchmarking large language models (LLMs) is critical for understanding their capabilities, limitations, and robustness. In addition to interface ar

model-releasesarxiv-cs-cl
26 Jun 2026
Safety

OPID: On-Policy Skill Distillation for Agentic Reinforcement Learning

DGX agent

arXiv:2606.26790v1 Announce Type: new Abstract: Outcome-based reinforcement learning provides a stable optimization backbone for language agents, but its sparse trajectory-level rewards provide little

safetyarxiv-cs-cl
26 Jun 2026
Research

Orthogonal Hierarchical Decomposition for Structure-Aware Table Understanding with Large Language Models

DGX agent

arXiv:2602.01969v2 Announce Type: replace Abstract: Complex tables with multi-level headers, merged cells and heterogeneous layouts pose persistent challenges for LLMs in both understanding and reason

researcharxiv-cs-cl
26 Jun 2026
Safety

Overcoming State Inertia: Minimally Invasive Temporal Alignment for Evolving Contexts

DGX agent

arXiv:2512.03704v3 Announce Type: replace Abstract: Long-context dialogue systems suffer from state inertia, where models over-attend to history and fail to adapt to evolving intents. We demonstrate t

safetyarxiv-cs-cl
26 Jun 2026
Safety

Paved with True Intents: Intent-Aware Training Improves LLM Safety Classification Across Training Regimes

DGX agent

arXiv:2606.27210v1 Announce Type: new Abstract: We argue that safety classifiers should model user intent as an explicit signal between the prompt and the final label. To study this, we introduce AIMS

safetyarxiv-cs-cl
26 Jun 2026
Research

Phonetic and semantic analyses of spoken corpora of Beijing and Taiwan Mandarin indicate that the neutral tone is a lexical tone

DGX agent

arXiv:2606.26360v1 Announce Type: new Abstract: The neutral, or floating, tone of Mandarin Chinese is a tone with an enigmatic set of properties. It has been described as a reduced tone, or as a tone

researcharxiv-cs-cl
26 Jun 2026
Local Ai

ProfileFoundry: A Synthetic Person-Object Substrate for Privacy, Memory, and Tool-Use Evaluation in LLM Agent

DGX agent

arXiv:2606.26403v1 Announce Type: new Abstract: Foundation-model research increasingly needs data about people: user state, personal histories, relationships, contact-like fields, documents, and longi

local-aiarxiv-cs-cl
26 Jun 2026
Agents

Prompt, Plan, Extract: Zero-Shot Agentic LLMs Workflows for Lung Pathology Extraction from Clinical Narratives

DGX agent

arXiv:2606.19852v2 Announce Type: replace Abstract: Information extraction from pathology reports is essential for cancer staging, tumor registry population. Yet key data remains embedded in narrative

agentsarxiv-cs-cl
26 Jun 2026
Model Releases

RedVox: Safety and Fairness Gaps in Speech Models Across Languages

DGX agent

arXiv:2606.26968v1 Announce Type: new Abstract: Speech-capable models are increasingly deployed in real-world applications across languages. Yet their safety and fairness beyond English settings and u

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Reproducibility Study of 'AlphaEdit: Null-Space Constrained Knowledge Editing for Language Models'

DGX agent

arXiv:2606.26783v1 Announce Type: cross Abstract: Fang et al. (2025) introduced a null-space constrained projection, named AlphaEdit, for locate-then-edit knowledge editing methods, theoretically guar

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

SocialPersona: Benchmarking Personalized Profiling and Response with Multimodal Social-Media Context

DGX agent

arXiv:2606.26654v1 Announce Type: new Abstract: Personalized language-model assistants are often evaluated through a memory lens: can a model recall preferences users have explicitly stated in dialogu

model-releasesarxiv-cs-cl
26 Jun 2026
Safety

Soft Token Alignment for Cross-Lingual Reasoning

DGX agent

arXiv:2606.26466v1 Announce Type: new Abstract: Multilingual large language models often produce inconsistent reasoning and answers for semantically equivalent prompts in different languages. Prior wo

safetyarxiv-cs-cl
26 Jun 2026
Safety

Somatic in the East, Psychological in the West?: Investigating Clinically-Grounded Cross-Cultural Depression Symptom Expression in LLMs

DGX agent

arXiv:2508.03247v2 Announce Type: replace Abstract: Prior clinical psychology research shows that Western individuals with depression tend to report psychological symptoms, while Eastern individuals r

safetyarxiv-cs-cl
26 Jun 2026
Model Releases

Spurious Rewards Paradox: Mechanistically Understanding How RLVR Activates Memorization Shortcuts in LLMs

DGX agent

arXiv:2601.11061v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is highly effective for enhancing LLM reasoning, yet recent evidence shows models like Q

model-releasesarxiv-cs-cl
26 Jun 2026
Safety

Staying VIGILant: Mitigating Visual Laziness via Counterfactual Visual Alignment in MLLMs

DGX agent

arXiv:2606.26387v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) extend large language models (LLMs) with visual perception, enabling joint reasoning over images and text. De

safetyarxiv-cs-cl
26 Jun 2026
Tutorials

Structure Before Collapse: Transient semantic geometry in next-token prediction

DGX agent

arXiv:2606.26749v1 Announce Type: cross Abstract: Neural Collapse predicts that balanced one-hot classification pushes model representations to be equally far from each other; a symmetric configuratio

tutorialsarxiv-cs-cl
26 Jun 2026
Research

Syntactic Belief Update as the Driver of Garden Path Processing Difficulty

DGX agent

arXiv:2606.27206v1 Announce Type: new Abstract: Garden path sentences present a processing difficulty for humans -- the sentence prefix leads the listener towards one interpretation, until the listene

researcharxiv-cs-cl
26 Jun 2026
Model Releases

Term-Centric Hierarchy Induction from Heterogeneous Corpora

DGX agent

arXiv:2606.26963v1 Announce Type: new Abstract: Organizing knowledge from diverse text sources into interpretable hierarchies is crucial for tasks such as policy analysis, innovation monitoring, and e

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

The Geometry of Updates: Fisher Alignment at Vocabulary Scale

DGX agent

arXiv:2606.27242v1 Announce Type: cross Abstract: Training-free source selection for LLM families with shared vocabularies arises in scientific string domains such as SMILES, protein, and genomic sequ

model-releasesarxiv-cs-cl
26 Jun 2026
Research

The Riddle Riddle: Testing Flexible Reasoning in Large Language Models and Humans

DGX agent

arXiv:2606.27103v1 Announce Type: new Abstract: Humans flexibly adapt their reasoning strategies to the requirements of a given problem. Large language models (LLMs) have performed well on many cognit

researcharxiv-cs-cl
26 Jun 2026
Model Releases

Towards Explainable Adjudicative Variance: Quantifying Judicial Discretion via Gated Multi-Task Learning

DGX agent

arXiv:2606.27069v1 Announce Type: new Abstract: Legal outcome prediction must disentangle objective case facts from adjudicative context. Merit-based rulings rely on factual evidence while technical d

model-releasesarxiv-cs-cl
26 Jun 2026
Research

Utilizing Cognitive Signals Generated during Human Reading to Enhance Keyphrase Extraction from Microblogs

DGX agent

arXiv:2606.26485v1 Announce Type: new Abstract: Microblogging platforms generate massive amounts of short, noisy, and dispersed user content, making automatic keyphrase extraction (AKE) an important b

researcharxiv-cs-cl
26 Jun 2026
Research

Vis-CoT: A Human-in-the-Loop Framework for Interactive Visualization and Intervention in LLM Chain-of-Thought Reasoning

DGX agent

arXiv:2509.01412v3 Announce Type: replace Abstract: Large language models (LLMs) show strong reasoning via chain-of-thought (CoT) prompting, but the process is opaque, which makes verification, debugg

researcharxiv-cs-cl
26 Jun 2026
Model Releases

When Actions Go Off-Task: Detecting and Correcting Misaligned Actions in Computer-Use Agents

DGX agent

arXiv:2602.08995v2 Announce Type: replace Abstract: Computer-use agents (CUAs) have made tremendous progress in the past year, yet they still frequently produce misaligned actions that deviate from th

model-releasesarxiv-cs-cl
26 Jun 2026
Research

Where Larger Models Excel: The Primacy of Constraint-Guided Reasoning

DGX agent

arXiv:2606.26108v1 Announce Type: new Abstract: Larger language models consistently outperform smaller ones on reasoning benchmarks, yet the reasoning differences underlying this gap remain underexplo

researcharxiv-cs-cl
26 Jun 2026
Research

Zero-shot Tweet-Level Stance Detection Enhanced by External Knowledge and Reflective Chain-of-Thought Reasoning

DGX agent

arXiv:2606.26571v1 Announce Type: new Abstract: Zero-shot tweet-level stance detection confronts two primary challenges: (1) mitigating the context sparsity inherent in short texts, and (2) establishi

researcharxiv-cs-cl
26 Jun 2026
Model Releases

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation

DGX agent

arXiv:2606.25476v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated remarkable performance across natural language processing tasks, yet their deployment in high-stakes appl

model-releasesarxiv-cs-cl
25 Jun 2026
Safety

A Survey of Toxicity Detection and Mitigation Strategies for Multilingual Language Models

DGX agent

arXiv:2606.25380v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed across languages, but their safety behavior remains uneven across linguistic and cultural context

safetyarxiv-cs-cl
25 Jun 2026
Research

A Systematic Analysis of Hybrid Linear Attention

DGX agent

arXiv:2507.06457v2 Announce Type: replace Abstract: Transformers face quadratic complexity and memory issues with long sequences, prompting the adoption of linear attention mechanisms using fixed-size

researcharxiv-cs-cl
25 Jun 2026
Research

Adapting Self-Supervised Speech Representations for Cross-lingual Dysarthria Detection in Parkinson's Disease

DGX agent

arXiv:2603.22225v3 Announce Type: replace Abstract: The limited availability of dysarthric speech data makes cross-lingual detection an important but challenging problem. A key difficulty is that spee

researcharxiv-cs-cl
25 Jun 2026
Safety

Adaptive Oscillatory Inductive Bias for Modeling Sharp Prosodic Dynamics in Diffusion-Based TTS

DGX agent

arXiv:2606.25424v1 Announce Type: cross Abstract: Diffusion-based text-to-speech (TTS) models have achieved significant improvements in speech quality. However, modeling sharp prosodic transitions and

safetyarxiv-cs-cl
25 Jun 2026
Agents

AgentOdyssey: Open-Ended Long-Horizon Text Game Generation for Test-Time Continual Learning Agents

DGX agent

arXiv:2606.24893v1 Announce Type: new Abstract: For agents to learn continuously from interaction with the world at test time, they must be able to explore effectively, acquire new world knowledge and

agentsarxiv-cs-cl
25 Jun 2026
Agents

AI translation of literary texts is 'fine', but readers still prefer human translations

DGX agent

arXiv:2606.26040v1 Announce Type: new Abstract: AI translation of literary works is increasingly common. While the content may be rendered adequately, we do not know enough about how readers experienc

agentsarxiv-cs-cl
25 Jun 2026
Safety

Am I More Pointwise or Pairwise? Revealing Position Bias in Rubric-Based LLM-as-a-Judge

DGX agent

arXiv:2602.02219v2 Announce Type: replace Abstract: Large language models are widely employed as evaluators, a paradigm commonly referred to as LLM-as-a-judge. Prior research has predominantly examine

safetyarxiv-cs-cl
25 Jun 2026
Research

An Empirical Study of Many-Shot In-Context Learning for Machine Translation of Low-Resource Languages

DGX agent

arXiv:2604.02596v3 Announce Type: replace Abstract: In-context learning (ICL) allows large language models (LLMs) to adapt to new tasks from a few examples, making it promising for languages underrepr

researcharxiv-cs-cl
25 Jun 2026
Research

Approximate Structured Diffusion for Sequence Labelling

DGX agent

arXiv:2606.18856v2 Announce Type: replace Abstract: Sequence labelling, a core task of Natural Language Processing (NLP), consists in assigning each token of an input sentence a label. From a Machine

researcharxiv-cs-cl
25 Jun 2026
Safety

ASAP: Agent-System Co-Design for Wall-Clock-Centered Auto HPO Research for ML Experiments

DGX agent

arXiv:2606.25207v1 Announce Type: cross Abstract: Hyperparameter Optimization (HPO) is essential for maximizing machine learning model performance, and its core challenge is sample efficiency: finding

safetyarxiv-cs-cl
25 Jun 2026
Agents

Autodata: An agentic data scientist to create high quality synthetic data

DGX agent

arXiv:2606.25996v1 Announce Type: cross Abstract: We introduce Autodata, a general method that enables AI agents to act as data scientists who build high quality training and evaluation data. We show

agentsarxiv-cs-cl
25 Jun 2026
Local Ai

Automatic Generation of Highlights for Academic Paper Via Prompt-based Learning

DGX agent

arXiv:2606.25253v1 Announce Type: new Abstract: Highlights provide a concise summary of the main contributions of an academic paper and help readers quickly understand its focus. However, many journal

local-aiarxiv-cs-cl
25 Jun 2026
Model Releases

Beyond Function Calling: Benchmarking Tool-Using Agents under Tool-Environment Unreliability

DGX agent

arXiv:2606.25819v1 Announce Type: new Abstract: Large language models are increasingly deployed as agents that solve tasks by interacting with external tool environments. Although recent tool-use benc

model-releasesarxiv-cs-cl
25 Jun 2026
Safety

Beyond Next-Observation Prediction: Agent-Authored World Modeling for Sequential Decision Making

DGX agent

arXiv:2606.25421v1 Announce Type: new Abstract: Recent studies on world modeling for Large Language Model (LLM) agents typically formulate the learning objective as next-observation prediction. Howeve

safetyarxiv-cs-cl
25 Jun 2026
Local Ai

BiPACE: Bisimulation-Guided Policy Optimization with Action Counterfactual Estimation for LLM Agents

DGX agent

arXiv:2606.25556v1 Announce Type: new Abstract: Stepwise group-based RL is an attractive way to train long-horizon LLM agents without a learned critic: it reuses multiple sampled rollouts to estimate

local-aiarxiv-cs-cl
25 Jun 2026
← Previous
1…4142434445…161
Next →