AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
Research

Token-weighted Direct Preference Optimization with Attention

DGX agent

arXiv:2605.21883v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) aligns Large Language Models with human preferences without the need for a separate reward model. However, DPO trea

researcharxiv-cs-cl
22 May 2026
Research

Tokenisation via Convex Relaxations

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2605.22821v1 Announce Type: new Abstract: Tokenisation is an integral part of the current NLP pipeline. Current tokenisation algorithms such as BPE and Unigram are greedy algorithms -- they make

researcharxiv-cs-cl
22 May 2026
Model Releases

Tokenization with Split Trees

DGX agent

arXiv:2605.22705v1 Announce Type: new Abstract: We introduce Tokenization with Split Trees (ToaST), a subword tokenization method that directly optimizes compression under a new recursive inference pr

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Training-Trajectory-Aware Token Selection

DGX agent

arXiv:2601.10348v2 Announce Type: replace Abstract: Efficient distillation is a key pathway for converting expensive reasoning capability into deployable efficiency, yet in the frontier regime where t

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation

DGX agent

arXiv:2605.22355v1 Announce Type: new Abstract: Public transit route planning traditionally depends on structured map infrastructure and complex routing engines, and no existing dataset supports train

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Two is better than one: A Collapse-free Multi-Reward RLIF Training Framework

DGX agent

arXiv:2605.22620v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has substantially improved the reasoning ability of LLMs, but often depends on external supervis

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Understanding Data Temporality Impact on Large Language Models Pre-training

DGX agent

arXiv:2605.22769v1 Announce Type: new Abstract: Large language models (LLMs) are typically trained on shuffled corpora, yielding models whose knowledge is frozen at train time and whose temporal groun

model-releasesarxiv-cs-cl
22 May 2026
Tutorials

Unified Data Selection for LLM Reasoning

DGX agent

arXiv:2605.22389v1 Announce Type: new Abstract: Effectively training Large Language Models (LLMs) for complex, long-CoT reasoning is often bottlenecked by the need for massive high-quality reasoning d

tutorialsarxiv-cs-cl
22 May 2026
Safety

Unifying Masked Diffusion Models with Various Generation Orders and Beyond

DGX agent

arXiv:2602.02112v2 Announce Type: replace-cross Abstract: Masked diffusion models (MDMs) are a potential alternative to autoregressive models (ARMs) for language generation, but generation quality dep

safetyarxiv-cs-cl
22 May 2026
Safety

UniSD: Towards a Unified Self-Distillation Framework for Large Language Models

DGX agent

arXiv:2605.06597v2 Announce Type: replace Abstract: Self-distillation (SD) offers a promising path for adapting large language models (LLMs) without relying on stronger external teachers. However, SD

safetyarxiv-cs-cl
22 May 2026
Safety

Value-Gradient Hypothesis of RL for LLMs

DGX agent

arXiv:2605.21654v1 Announce Type: cross Abstract: Reinforcement learning substantially improves pretrained language models, but it remains understudied why critic-free methods such as PPO and GRPO wor

safetyarxiv-cs-cl
22 May 2026
Safety

Vector Policy Optimization: Training for Diversity Improves Test-Time Search

DGX agent

arXiv:2605.22817v1 Announce Type: cross Abstract: Language models must now generalize out of the box to novel environments and work inside inference-scaling search procedures, such as AlphaEvolve, tha

safetyarxiv-cs-cl
22 May 2026
Model Releases

When Cases Get Rare: A Retrieval Benchmark for Off-Guideline Clinical Question Answering

DGX agent

arXiv:2605.21807v1 Announce Type: new Abstract: Across medical specialties, clinical practice is anchored in evidence-based guidelines that codify best studied diagnostic and treatment pathways. These

model-releasesarxiv-cs-cl
22 May 2026
Research

When Shared Knowledge Hurts: Spectral Over-Accumulation in Model Merging

DGX agent

arXiv:2602.05536v2 Announce Type: replace-cross Abstract: Model merging combines multiple fine-tuned models into a single model by adding their weight updates, providing a lightweight alternative to r

researcharxiv-cs-cl
22 May 2026
Research

Whose Voice Counts? Mapping Stakeholder Perspectives on AI Through Public Submissions to the U.S. Government

DGX agent

arXiv:2605.22650v1 Announce Type: new Abstract: As artificial intelligence (AI) systems become more common in our daily lives, it is important to understand how different stakeholders comprehend and e

researcharxiv-cs-cl
22 May 2026
Safety

Why Semantic Entropy Fails: Geometry-Aware and Calibrated Uncertainty for Policy Optimization

DGX agent

arXiv:2605.21801v1 Announce Type: cross Abstract: Post-training has become central to improving reasoning and alignment in large language models, where critic-free models enable scalable learning from

safetyarxiv-cs-cl
22 May 2026
Safety

'Would You Want an AI Tutor?' Understanding Stakeholder Perceptions of LLM-based Systems in the Classroom

DGX agent

arXiv:2503.02885v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have gained traction in educational settings, often framed as virtual tutors or teaching assistants. Following ea

safetyarxiv-cs-cl
22 May 2026
Model Releases

X-Token: Projection-Guided Cross-Tokenizer Knowledge Distillation

DGX agent

arXiv:2605.21699v1 Announce Type: cross Abstract: Cross-tokenizer knowledge distillation allows a student model to learn from teachers with incompatible vocabularies. Prior work operates on hidden sta

model-releasesarxiv-cs-cl
22 May 2026
Safety

A Systematic Comparison between Extractive Self-Explanations and Human Rationales in Text Classification

DGX agent

arXiv:2410.03296v4 Announce Type: replace Abstract: Instruction-tuned LLMs are able to provide extit{an} explanation about their output to users by generating self-explanations, without requiring the

safetyarxiv-cs-cl
21 May 2026
Model Releases

ACL-Verbatim: hallucination-free question answering for research

DGX agent

arXiv:2605.21102v1 Announce Type: new Abstract: Academic researchers need efficient and reliable methods for collecting high-quality information from trusted sources, but modern tools for AI-assisted

model-releasesarxiv-cs-cl
21 May 2026
Safety

AFD-INSTRUCTION: A Comprehensive Antibody Instruction Dataset with Functional Annotations for LLM-Based Understanding and Design

DGX agent

arXiv:2602.04916v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have significantly advanced protein representation learning. However, their capacity to interpret and design anti

safetyarxiv-cs-cl
21 May 2026
Model Releases

AgentAtlas: Beyond Outcome Leaderboards for LLM Agents

DGX agent

arXiv:2605.20530v1 Announce Type: cross Abstract: Large language model agents now act on codebases, browsers, operating systems, calendars, files, and tool ecosystems, but the benchmarks used to evalu

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

AI-Assisted Scientific Assessment: A Case Study on Climate Change

DGX agent

arXiv:2602.09723v2 Announce Type: replace Abstract: The emerging paradigm of AI co-scientists focuses on tasks characterized by repeatable verification, where agents explore search spaces in 'guess an

model-releasesarxiv-cs-cl
21 May 2026
Research

AI-Augmented Surveys: Leveraging Large Language Models and Surveys for Opinion Prediction

DGX agent

arXiv:2305.09620v4 Announce Type: replace Abstract: Nationally representative surveys track public opinion, yet they ask only a limited set of questions each year, limiting its potential to capture hi

researcharxiv-cs-cl
21 May 2026
Agents

AiraXiv: An AI-Driven Open-Access Platform for Human and AI Scientists

DGX agent

arXiv:2605.21481v1 Announce Type: cross Abstract: Recent advances in artificial intelligence (AI) have accelerated the growth of both human-authored and AI-generated research outputs, placing increasi

agentsarxiv-cs-cl
21 May 2026
Model Releases

Anatomy of Agentic Memory: Taxonomy and Empirical Analysis of Evaluation and System Limitations

DGX agent

arXiv:2602.19320v2 Announce Type: replace Abstract: Agentic memory systems enable large language model (LLM) agents to maintain state across long interactions, supporting long-horizon reasoning and pe

model-releasesarxiv-cs-cl
21 May 2026
Applications

Anti-establishment sentiment on TikTok: Implications for understanding influence(rs) and expertise on social media

DGX agent

arXiv:2508.16453v2 Announce Type: replace-cross Abstract: Distrust of public serving institutions and anti-establishment views are on the rise (especially in the U.S.). As people turn to social media

applicationsarxiv-cs-cl
21 May 2026
Model Releases

APM: Evaluating Style Personalization in LLMs with Arbitrary Preference Mappings

DGX agent

arXiv:2605.21063v1 Announce Type: new Abstract: Typical LLM responses tend to follow a default style, even though users often have distinct preferences regarding tone, verbosity, and formality that th

model-releasesarxiv-cs-cl
21 May 2026
Research

ArPoMeme: An Annotated Arabic Multimodal Dataset for Political Ideology and Polarization

DGX agent

arXiv:2605.20967v1 Announce Type: new Abstract: Memes have become a prominent medium of political communication in the Arab world, reflecting how humor, imagery, and text interact to express ideologic

researcharxiv-cs-cl
21 May 2026
Research

Assessing socio-economic climate impacts from text data

DGX agent

arXiv:2605.20793v1 Announce Type: new Abstract: Recent advances in natural language processing (NLP) and large language models (LLMs) have enabled the systematic use of large-scale textual data from n

researcharxiv-cs-cl
21 May 2026
Agents

Auto-Dreamer: Learning Offline Memory Consolidation for Language Agents

DGX agent

arXiv:2605.20616v1 Announce Type: new Abstract: Language agents increasingly operate over streams of related tasks, yet existing memory systems struggle to convert accumulated experience into reusable

agentsarxiv-cs-cl
21 May 2026
Model Releases

Automated ICD Classification of Psychiatric Diagnoses: From Classical NLP to Large Language Models

DGX agent

arXiv:2605.21154v1 Announce Type: new Abstract: Mental health has become a global priority, leading to a massive administrative burden in the coding of clinical diagnoses. This study proposes the auto

model-releasesarxiv-cs-cl
21 May 2026
Safety

Automatically Learning Construction Injury Precursors from Text

DGX agent

arXiv:1907.11769v4 Announce Type: replace Abstract: In light of the increasing availability of digitally recorded safety reports in the construction industry, it is important to develop methods to exp

safetyarxiv-cs-cl
21 May 2026
Safety

AVSD: Adaptive-View Self-Distillation by Balancing Consensus and Teacher-Specific Privileged Signals

DGX agent

arXiv:2605.20643v1 Announce Type: cross Abstract: Self-distillation enables language models to learn on-policy from their own trajectories by using the same model as both student and teacher, with the

safetyarxiv-cs-cl
21 May 2026
Safety

Bayesian Preference Learning for Test-Time Steerable Reward Models

DGX agent

arXiv:2602.08819v2 Announce Type: replace-cross Abstract: Reward models are central to aligning language models with human preferences via reinforcement learning (RL). As RL is increasingly applied to

safetyarxiv-cs-cl
21 May 2026
Applications

Beyond Semantic Similarity: A Two-Phase Non-Parametric Retrieval Workflow for Corporate Credit Underwriting

DGX agent

arXiv:2605.20684v1 Announce Type: new Abstract: Corporate credit underwriting requires analysts to extract actionable evidence from long, heterogeneous financial documents spanning hundreds of pages a

applicationsarxiv-cs-cl
21 May 2026
Safety

Beyond Text-to-SQL: An Agentic LLM System for Governed Enterprise Analytics APIs

DGX agent

arXiv:2605.21027v1 Announce Type: new Abstract: Enterprise analytics aims to make organizational data accessible for decision-making, yet non-technical users still face barriers when using traditional

safetyarxiv-cs-cl
21 May 2026
Agents

Beyond Words: Multimodal LLM Knows When to Speak

DGX agent

arXiv:2505.14654v2 Announce Type: replace-cross Abstract: Chatbots via large language models (LLMs) generate fluent responses but often struggle with when to speak, especially for brief, timely listen

agentsarxiv-cs-cl
21 May 2026
Applications

Bridging Language Models and Financial Analysis

DGX agent

arXiv:2503.22693v2 Announce Type: replace-cross Abstract: The rapid advancements in Large Language Models (LLMs) have unlocked transformative possibilities in natural language processing, particularly

applicationsarxiv-cs-cl
21 May 2026
Tutorials

Building a Custom Taxonomy of AI Skills and Tasks from the Ground Up with Job Postings

DGX agent

arXiv:2605.21029v1 Announce Type: new Abstract: Utilizing LLMs for automated taxonomy construction presents a clear opportunity for the comprehensive, yet efficient mapping of potentially complex doma

tutorialsarxiv-cs-cl
21 May 2026
Research

Building Arabic NLP from the Ground Up: Twenty Years of Lessons, Failures, and Open Problems

DGX agent

arXiv:2605.20786v1 Announce Type: new Abstract: This paper reflects on twenty years of building NLP resources and research infrastructure for Arabic, a language spoken by hundreds of millions yet hist

researcharxiv-cs-cl
21 May 2026
Model Releases

Calibration vs Decision Making: Revisiting the Reliability Paradox in Unlearned Language Models

DGX agent

arXiv:2605.20915v1 Announce Type: new Abstract: Machine unlearning aims to remove the influence of specific training data from a model while preserving reliable behavior on the remaining data, making

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Causal Path Alignment: Anchoring the Optimization Trajectory for Controllable In-Parameter Knowledge Editing

DGX agent

arXiv:2506.04042v2 Announce Type: replace Abstract: Knowledge editing is pivotal for efficiently updating the parametric memory of Large Language Models (LLMs), enabling them to function as evolving a

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Chronicle: A Multimodal Foundation Model for Joint Language and Time Series Understanding

DGX agent

arXiv:2605.20268v1 Announce Type: cross Abstract: Real-world time series come with text: metadata, descriptions, news, reports. Yet time series foundation models process numerical sequences in isolati

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

ChunkFT: Byte-Streamed Optimization for Memory-Efficient Full Fine-Tuning

DGX agent

arXiv:2605.21177v1 Announce Type: cross Abstract: This work presents extsc{ChunkFT}, a memory-efficient fine-tuning framework that reformulates full-parameter fine-tuning around a dynamically activate

model-releasesarxiv-cs-cl
21 May 2026
Tutorials

Collocational bootstrapping: A hypothesis about the learning of subject-verb agreement in humans and neural networks

DGX agent

arXiv:2605.20529v1 Announce Type: new Abstract: In what ways might statistical signals in linguistic input assist with the acquisition of syntax? Here we hypothesize a mechanism called collocational b

tutorialsarxiv-cs-cl
21 May 2026
Model Releases

CP-MoE: Consistency-Preserving Mixture-of-Experts for Continual Learning

DGX agent

arXiv:2605.20247v1 Announce Type: cross Abstract: Catastrophic forgetting remains a major obstacle to continual learning in large language models (LLMs) and vision--language models (VLMs). Although Mi

model-releasesarxiv-cs-cl
21 May 2026
Safety

Cross-lingual robustness of LLM-brain alignment and its computational roots

DGX agent

arXiv:2605.21049v1 Announce Type: new Abstract: Large language models (LLMs) reliably predict neural activity during language comprehension and transformer depth has been interpreted as mirroring hier

safetyarxiv-cs-cl
21 May 2026
← Previous
1…8384858687…162
Next →