AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

The Position Curse: LLMs Struggle to Locate the Last Few Items in a List

DGX agent

arXiv:2605.07127v1 Announce Type: cross Abstract: Modern large language models (LLMs) can find a needle in a haystack (locating a single relevant fact buried among hundreds of thousands of irrelevant

model-releasesarxiv-cs-cl
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

The Proxy Presumption: From Semantic Embeddings to Valid Social Measures

DGX agent

arXiv:2605.07409v1 Announce Type: new Abstract: Natural Language Processing is rapidly evolving into a primary instrument for Computational Social Science, with researchers increasingly using embeddin

safetyarxiv-cs-cl
11 May 2026
Safety

Theoretical Limits of Language Model Alignment

DGX agent

arXiv:2605.07105v1 Announce Type: cross Abstract: Language model (LM) alignment improves model outputs to reflect human preferences while preserving the capabilities of the base model. The most common

safetyarxiv-cs-cl
11 May 2026
Safety

Think-with-Rubrics: From External Evaluator to Internal Reasoning Guidance

DGX agent

arXiv:2605.07461v1 Announce Type: new Abstract: Rubrics have been extensively utilized for evaluating unverifiable, open-ended tasks, with recent research incorporating them into reward systems for re

safetyarxiv-cs-cl
11 May 2026
Model Releases

Topic Is Not Agenda: A Citation-Community Audit of Text Embeddings

DGX agent

arXiv:2605.07158v1 Announce Type: cross Abstract: Vector search and retrieval-augmented generation (RAG) rest on the assumption that cosine similarity between text embeddings reflects conceptual relat

model-releasesarxiv-cs-cl
11 May 2026
Local Ai

Topology-Enhanced Alignment for Large Language Models: Trajectory Topology Loss and Topological Preference Optimization

DGX agent

arXiv:2605.07172v1 Announce Type: new Abstract: Alignment of large language models (LLMs) via SFT and RLHF/DPO typically ignores the global geometry of the representation space, relying instead on loc

local-aiarxiv-cs-cl
11 May 2026
Model Releases

Towards Closing the Autoregressive Gap in Language Modeling via Entropy-Gated Continuous Bitstream Diffusion

DGX agent

arXiv:2605.07013v1 Announce Type: new Abstract: Diffusion language models (DLMs) promise parallel, order-agnostic generation, but on standard benchmarks they have historically lagged behind autoregres

model-releasesarxiv-cs-cl
11 May 2026
Research

Tracing the Arrow of Time: Diagnosing Temporal Information Flow in Video-LLMs

DGX agent

arXiv:2605.07568v1 Announce Type: cross Abstract: The Arrow-of-Time (AoT) task, determining whether a video plays forward or backward by recognizing temporal irreversibility, is one humans solve with

researcharxiv-cs-cl
11 May 2026
Safety

Training-Free Multimodal Large Language Model Orchestration

DGX agent

arXiv:2508.10016v3 Announce Type: replace Abstract: Building interactive omni-modal assistants often relies on end-to-end multimodal alignment to fuse heterogeneous modalities, which incurs substantia

safetyarxiv-cs-cl
11 May 2026
Safety

UFT: Unifying Fine-Tuning of SFT and RLHF/DPO/UNA through a Generalized Implicit Reward Function

DGX agent

arXiv:2410.21438v3 Announce Type: replace Abstract: By pretraining on trillions of tokens, an LLM gains the capability of text generation. However, to enhance its utility and reduce potential harm, SF

safetyarxiv-cs-cl
11 May 2026
Safety

UNA: A Unified Supervised Framework for Efficient LLM Alignment Across Feedback Types

DGX agent

arXiv:2408.15339v4 Announce Type: replace-cross Abstract: RL alignment methods, including RLHF and DPO, are primarily based on pairwise preference data. Although scalar or score-based feedback has bee

safetyarxiv-cs-cl
11 May 2026
Research

Uncertainty-Aware Structured Data Extraction from Full CMR Reports via Distilled LLMs

DGX agent

arXiv:2605.08045v1 Announce Type: new Abstract: Converting free-text cardiac magnetic resonance (CMR) reports into auditable structured data remains a bottleneck for cohort assembly, longitudinal cura

researcharxiv-cs-cl
11 May 2026
Applications

User eXperience Perception Insights Dataset (UXPID): Synthetic User Feedback from Public Industrial Forums

DGX agent

arXiv:2509.11777v2 Announce Type: replace Abstract: Customer feedback in industrial forums offers rich but underexplored insights into real-world product experience. Yet systematic analysis remains ch

applicationsarxiv-cs-cl
11 May 2026
Model Releases

Utility-Preserving De-Identification for Math Tutoring: Investigating Numeric Ambiguity in the MathEd-PII Benchmark Dataset

DGX agent

arXiv:2602.16571v2 Announce Type: replace Abstract: Large-scale sharing of dialogue data is key to advancing the science of teaching and learning, yet rigorous de-identification remains a major barrie

model-releasesarxiv-cs-cl
11 May 2026
Research

WeatherSyn: An Instruction Tuning MLLM For Weather Forecasting Report Generation

DGX agent

arXiv:2605.07522v1 Announce Type: new Abstract: Accurate weather forecast reporting enables individuals and communities to better plan daily activities and agricultural operations. However, the curren

researcharxiv-cs-cl
11 May 2026
Model Releases

When Are Experts Misrouted? Counterfactual Routing Analysis in Mixture-of-Experts Language Models

DGX agent

arXiv:2605.07260v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) language models route each token to a small subset of experts, but whether the routes selected by a trained top-k router are

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

When Routine Chats Turn Toxic: Unintended Long-Term State Poisoning in Personalized Agents

DGX agent

arXiv:2605.06731v1 Announce Type: cross Abstract: Personalized LLM agents maintain persistent cross-session state to support long-horizon collaboration. Yet, this persistence introduces a subtle but c

model-releasesarxiv-cs-cl
11 May 2026
Research

Why do Large Language Models Fail in Low-resource Translation? Unraveling the Token Dynamics of Large Language Models for Machine Translation

DGX agent

arXiv:2605.07533v1 Announce Type: new Abstract: Large Language Models (LLMs) have recently demonstrated strong performance in machine translation (MT). However, most prior work focuses on improving or

researcharxiv-cs-cl
11 May 2026
Research

WorldCup Sampling for Multi-bit LLM Watermarking

DGX agent

arXiv:2602.01752v2 Announce Type: replace Abstract: As large language models (LLMs) generate increasingly human-like text, watermarking has emerged as a promising solution for reliable attribution bey

researcharxiv-cs-cl
11 May 2026
Applications

A Comparative Analysis of Machine Learning and Deep Learning Models for Tweet Sentiment Classification: A Case Study on the Sentiment140 Dataset

DGX agent

arXiv:2605.04888v1 Announce Type: new Abstract: The exponential growth of social media has created an urgent need for automated systems to analyze unstructured public sentiment in real time. This stud

applicationsarxiv-cs-cl
7 May 2026
Model Releases

A Comparative Study of PyCaret AutoML and CNN-BiLSTM for Binary Hate Speech Detection in Indonesian Twitter

DGX agent

arXiv:2605.04885v1 Announce Type: new Abstract: This paper compares a PyCaret AutoML branch and a CNN-BiLSTM branch for binary hate speech detection on Indonesian Twitter using the HS label from the c

model-releasesarxiv-cs-cl
7 May 2026
Applications

A Hybrid Method for Low-Resource Named Entity Recognition

DGX agent

arXiv:2605.04489v1 Announce Type: cross Abstract: Named Entity Recognition (NER) is a critical component of Natural Language Processing with diverse applications in information extraction and conversa

applicationsarxiv-cs-cl
7 May 2026
Safety

Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning

DGX agent

arXiv:2605.04066v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is an essential paradigm that enhances the reasoning capabilities of Large Language Models (LLMs).

safetyarxiv-cs-cl
7 May 2026
Model Releases

Adapting Large Language Models to a Low-Resource Agglutinative Language: A Comparative Study of LoRA and QLoRA for Bashkir

DGX agent

arXiv:2605.04948v1 Announce Type: new Abstract: This paper presents a comparative study of parameter-efficient fine-tuning (PEFT) methods, including LoRA and QLoRA, applied to the task of adapting lar

model-releasesarxiv-cs-cl
7 May 2026
Safety

Anticipating Innovation Using Large Language Models

DGX agent

arXiv:2605.04875v1 Announce Type: new Abstract: Forecasting innovation, intended as the emergence of new technological combinations, is a fundamental challenge for science and policy. We show that for

safetyarxiv-cs-cl
7 May 2026
Model Releases

Are LLMs Ready for Conflict Monitoring? Empirical Evidence from West Africa

DGX agent

arXiv:2605.04177v1 Announce Type: new Abstract: As LLMs enter conflict monitoring, understanding systematic distortions in their outputs is critical for humanitarian accountability. We evaluate four v

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Assessing Cognitive Effort in L2 Idiomatic Processing: An Eye-Tracking Dataset

DGX agent

arXiv:2605.04857v1 Announce Type: new Abstract: This paper presents the development and validation of an eye-tracking dataset designed to investigate how second-language (L2) learners process idiomati

model-releasesarxiv-cs-cl
7 May 2026
Applications

Automatically Finding and Validating Unexpected Side-Effects of Interventions on Language Models

DGX agent

arXiv:2605.05090v1 Announce Type: new Abstract: We present an automated, contrastive evaluation pipeline for auditing the behavioral impact of interventions on large language models. Given a base mode

applicationsarxiv-cs-cl
7 May 2026
Research

Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation

DGX agent

arXiv:2605.04128v1 Announce Type: cross Abstract: We present JoyAI-Image, a unified multimodal foundation model for visual understanding, text-to-image generation, and instruction-guided image editing

researcharxiv-cs-cl
7 May 2026
Safety

Balanced Aggregation: Understanding and Fixing Aggregation Bias in GRPO

DGX agent

arXiv:2605.04077v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a central paradigm for improving reasoning and code generation in large language mode

safetyarxiv-cs-cl
7 May 2026
Model Releases

Benchmarking POS Tagging for the Tajik Language: A Comparative Study of Neural Architectures on the TajPersParallel Corpus

DGX agent

arXiv:2605.04576v1 Announce Type: new Abstract: This paper presents the first benchmark for the task of automatic part-of-speech (POS) tagging for the Tajik language. Despite the existence of multilin

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

BenCSSmark: Making the Social Sciences Count in LLM Research

DGX agent

arXiv:2605.04886v1 Announce Type: new Abstract: This position paper argues that the under-representation of social science tasks in contemporary LLM benchmarks limits advances in both LLM evaluation a

model-releasesarxiv-cs-cl
7 May 2026
Research

Beyond Exponential Decay: Rethinking Error Accumulation in Large Language Models

DGX agent

arXiv:2505.24187v2 Announce Type: replace Abstract: The prevailing assumption of an exponential decay in large language model (LLM) reliability with sequence length, predicated on independent per-toke

researcharxiv-cs-cl
7 May 2026
Safety

Beyond Public Access in LLM Pre-Training Data

DGX agent

arXiv:2505.00020v2 Announce Type: replace Abstract: Using a legally obtained dataset of 34 copyrighted O'Reilly Media books, we apply the DE-COP membership inference attack method to investigate wheth

safetyarxiv-cs-cl
7 May 2026
Applications

Beyond Semantics: An Evidential Reasoning-Aware Multi-View Learning Framework for Trustworthy Mental Health Prediction

DGX agent

arXiv:2605.05121v1 Announce Type: new Abstract: Automated mental health prediction using textual data has shown promising results with deep learning and large language models. However, deploying these

applicationsarxiv-cs-cl
7 May 2026
Model Releases

Budgeted LoRA: Distillation as Structured Compute Allocation for Efficient Inference

DGX agent

arXiv:2605.04341v1 Announce Type: cross Abstract: We study distillation for large language models under explicit compute constraints, with the goal of producing student models that are not only cheape

model-releasesarxiv-cs-cl
7 May 2026
Research

CAR: Query-Guided Confidence-Aware Reranking for Retrieval-Augmented Generation

DGX agent

arXiv:2605.04495v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) depends on document ranking to provide useful evidence for generation, but conventional reranking methods mainly op

researcharxiv-cs-cl
7 May 2026
Safety

CHE-TKG: Collaborative Historical Evidence and Evolutionary Dynamics Learning for Temporal Knowledge Graph Reasoning

DGX agent

arXiv:2605.04652v1 Announce Type: new Abstract: Temporal knowledge graph (TKG) reasoning aims to predict future events from historical facts. A key challenge lies in jointly capturing two sources of p

safetyarxiv-cs-cl
7 May 2026
Model Releases

Conceptors for Semantic Steering

DGX agent

arXiv:2605.04980v1 Announce Type: cross Abstract: Activation-based steering provides control of LLM behavior at inference time, but the dominant paradigm reduces each concept to a single direction who

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Conflict-Aware Fusion: Mitigating Logic Inertia in Large Language Models via Structured Cognitive Priors

DGX agent

arXiv:2512.06393v5 Announce Type: replace-cross Abstract: Large language models (LLMs) achieve high accuracy on many reasoning benchmarks but remain brittle under structural perturbations of rule-base

model-releasesarxiv-cs-cl
7 May 2026
Safety

Connecting online criminal behavior with machine learning: Using authorship attribution to analyze and link potential online traffickers

DGX agent

arXiv:2605.04080v1 Announce Type: new Abstract: This research investigated how online criminal activities can be better understood and connected using data-driven machine learning methods. Many illega

safetyarxiv-cs-cl
7 May 2026
Research

Continual Knowledge Updating in LLM Systems: Learning Through Multi-Timescale Memory Dynamics

DGX agent

arXiv:2605.05097v1 Announce Type: cross Abstract: LLMs are trained once, then deployed into a world that never stops changing. External memory compensates for this, but most systems manage it explicit

researcharxiv-cs-cl
7 May 2026
Hardware

Coral: Cost-Efficient Multi-LLM Serving over Heterogeneous Cloud GPUs

DGX agent

arXiv:2605.04357v1 Announce Type: cross Abstract: The usage of large language models (LLMs) has grown increasingly fragmented, with no single model dominating. Meanwhile, cloud providers offer a wide

hardwarearxiv-cs-cl
7 May 2026
Research

Cross-Tokenizer Likelihood Scoring Algorithms for Language Model Distillation

DGX agent

arXiv:2512.14954v2 Announce Type: replace Abstract: Computing next-token likelihood ratios between two language models (LMs) is a standard task in training paradigms such as knowledge distillation. Si

researcharxiv-cs-cl
7 May 2026
Research

Detecting Hallucinations in Large Language Models via Internal Attention Divergence Signals

DGX agent

arXiv:2605.05025v1 Announce Type: new Abstract: We propose a lightweight and single-pass uncertainty quantification method for detecting hallucinations in Large Language Models. The method uses attent

researcharxiv-cs-cl
7 May 2026
Safety

DFPO: Scaling Value Modeling via Distributional Flow towards Robust and Generalizable LLM Post-Training

DGX agent

arXiv:2602.05890v2 Announce Type: replace-cross Abstract: Training reinforcement learning (RL) systems in real-world environments remains challenging due to noisy supervision and poor out-of-domain (O

safetyarxiv-cs-cl
7 May 2026
Research

DIAL: Direct Iterative Adversarial Learning for Realistic Multi-Turn Dialogue Simulation

DGX agent

arXiv:2512.20773v4 Announce Type: replace Abstract: Realistic user simulation is crucial for training and evaluating multi-turn dialogue systems, yet creating simulators that accurately replicate huma

researcharxiv-cs-cl
7 May 2026
Tutorials

Diffusion-Inspired Masked Fine-Tuning for Knowledge Injection in Autoregressive LLMs

DGX agent

arXiv:2510.09885v5 Announce Type: replace Abstract: Large language models (LLMs) are often used in environments where facts evolve, yet factual knowledge updates via fine-tuning on unstructured text o

tutorialsarxiv-cs-cl
7 May 2026
← Previous
1…103104105106107…161
Next →