AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Safety

IACM-RL: Intent-Aware Context Management and Reinforcement Learning for Complex Tool Invocation under Dynamic Intent Fluctuations

DGX agent

arXiv:2608.02110v1 Announce Type: new Abstract: Executing long-horizon tool invocations in real-world environments is severely challenged by dynamic user intent noise. Existing methods attempt robustn

safetyarxiv-cs-cl
4 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Illuminating Visual Identity in Universal Multimodal Embeddings

DGX agent

arXiv:2608.01794v1 Announce Type: cross Abstract: Universal Multimodal Embeddings (UMEs) aim to unify various modalities and tasks into a shared representation space. In recent years, this field has w

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

Instruction-Conditioned Exploration with Asymmetric Reinforcement Learning and Self-Distillation

DGX agent

arXiv:2608.02087v1 Announce Type: cross Abstract: Post-training Large Language Models (LLMs) with Reinforcement Learning (RL) has become an important tool for improving model capabilities, but the LLM

safetyarxiv-cs-cl
4 Aug 2026
Agents

Intern-S1-MO: Long-horizon Reasoning Agent for Olympiad?Level Mathematical Problem Solving

DGX agent

arXiv:2512.10739v3 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have expanded the mathematical reasoning frontier through Chain-of-Thought (CoT) techniques and Reinforcement Learning

agentsarxiv-cs-cl
4 Aug 2026
Research

Interpretable Recognition of Cognitive Distortions in Natural Language Texts

DGX agent

arXiv:2511.05969v2 Announce Type: replace Abstract: We propose a new approach to multi-factor classification of natural language texts based on weighted structured patterns such as N-grams, taking int

researcharxiv-cs-cl
4 Aug 2026
Research

Just on Time: Token-Level Early Stopping for Diffusion Language Models

DGX agent

arXiv:2602.11133v2 Announce Type: replace-cross Abstract: Diffusion language models generate text through iterative refinement, a process that is often computationally inefficient because many tokens

researcharxiv-cs-cl
4 Aug 2026
Model Releases

LangFIR: Discovering Sparse Language-Specific Features from Monolingual Data for Language Steering

DGX agent

arXiv:2604.03532v2 Announce Type: replace Abstract: Large language models (LLMs) show strong multilingual capabilities, yet reliably controlling the language of their outputs remains difficult. Repres

model-releasesarxiv-cs-cl
4 Aug 2026
Local Ai

Language Equality has a Price: A Systematic Investigation of Multi-turn LLM Performance for EU-24+

DGX agent

arXiv:2608.01395v1 Announce Type: new Abstract: We evaluate large language models (LLMs) as language agents playing goal-directed dialogue games in self-play across 30 languages: the 24 official EU la

local-aiarxiv-cs-cl
4 Aug 2026
Research

Leak It: A Probabilistic Approach to Training-Data Extraction from Black-Box Language Models

DGX agent

arXiv:2608.00144v1 Announce Type: cross Abstract: Membership inference (MIA) on language models is usually summarised by an aggregate ROC-AUC, but such evaluations are confounded: model-free blind bas

researcharxiv-cs-cl
4 Aug 2026
Model Releases

Learning What to Remember: Test-Time Training via Context Distillation

DGX agent

arXiv:2608.01672v1 Announce Type: new Abstract: Effective long-context modeling is not merely about retaining more of the past, but about preserving the information that may prove relevant later. Test

model-releasesarxiv-cs-cl
4 Aug 2026
Research

Length Penalties Make Chain-of-Thought Less Monitorable

DGX agent

arXiv:2607.09786v3 Announce Type: replace-cross Abstract: To curb overthinking and reduce inference costs, researchers now train reasoning models with penalties on chain of thought length. We find tha

researcharxiv-cs-cl
4 Aug 2026
Research

Leveraging Synthetic Data for Question Answering with Multilingual LLMs in the Agricultural Domain

DGX agent

arXiv:2507.16974v3 Announce Type: replace Abstract: Enabling farmers to access accurate agriculture-related information in their native languages in a timely manner is crucial for the success of the a

researcharxiv-cs-cl
4 Aug 2026
Model Releases

LiveMem: Maintaining Memory State Continuity in Long-Running LLM Inference

DGX agent

arXiv:2608.02515v1 Announce Type: new Abstract: Long-running assistants and agents consume interaction streams that eventually outgrow the context. Existing context retention, summarization, and retri

model-releasesarxiv-cs-cl
4 Aug 2026
Research

LLM generation novelty through the lens of semantic similarity

DGX agent

arXiv:2510.27313v3 Announce Type: replace-cross Abstract: Generation novelty is a key indicator of an LLM's ability to generalize, yet measuring it against full pretraining corpora is computationally

researcharxiv-cs-cl
4 Aug 2026
Safety

LLM-OSDA: An Optimal-Stopping Dynamic Auction for Native Advertising in Multi-Turn LLM Conversations

DGX agent

arXiv:2608.00123v1 Announce Type: new Abstract: LLM-native advertising embeds sponsored content directly into model-generated responses, shifting the unit of sale from a fixed slot to a moment within

safetyarxiv-cs-cl
4 Aug 2026
Safety

LM-mixup: Text Data Augmentation via Language Model based Mixup

DGX agent

arXiv:2510.20449v2 Announce Type: replace Abstract: Instruction tuning is crucial for aligning Large Language Models (LLMs), yet the quality of instruction-following data varies significantly. While h

safetyarxiv-cs-cl
4 Aug 2026
Research

Loanword or Switch? The Annotation Boundary, Not the Model, Drives Kazakh-Russian Code-Switching Identification

DGX agent

arXiv:2608.00581v1 Announce Type: new Abstract: Off-the-shelf LID and letter heuristics over-label Kazakh-Russian social text as mixed: Russian loanwords inside Kazakh look like code-switching under a

researcharxiv-cs-cl
4 Aug 2026
Model Releases

Long-Horizon Embodied Decision-Making via Multimodal Memory Compression

DGX agent

arXiv:2608.01456v1 Announce Type: cross Abstract: Agents are increasingly expected to act not only as task executors, but also as decision-makers on behalf of human users. This shift requires agents t

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

LongCat Sparse Attention: Taming the Lightning via Streaming-aware Hierarchical Cross-Layer Indexing

DGX agent

arXiv:2608.01662v1 Announce Type: cross Abstract: DeepSeek Sparse Attention (DSA) enables efficient long-context modeling through its Lightning Indexer. However, practical deployment remains constrain

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

LongChart VQA: A Comprehensive Benchmark for MLLMs with Complex Multi-Chart Reasoning

DGX agent

arXiv:2608.01328v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are rapidly evolving with expanded context windows and stronger reasoning capabilities, enabling multi-chart un

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

Look Ahead Before You Distill: Future Trajectory Validation of Teacher Guidance for Agentic On-Policy Distillation

DGX agent

arXiv:2608.01953v1 Announce Type: new Abstract: On-policy distillation (OPD) provides teacher supervision on states visited by the student, reducing the distribution gap between training and inference

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

LoopsBench: From Harness Engineering to Loop Engineering in Benchmarking Coding Agent

DGX agent

arXiv:2608.00267v1 Announce Type: cross Abstract: Coding agent infrastructure is shifting from harness engineering toward loop engineering as coding agents are deployed for sustained long-horizon soft

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

MAPLE: Metadata Augmented Private Language Evolution

DGX agent

arXiv:2603.19258v2 Announce Type: replace Abstract: Differentially private (DP) fine-tuning of large language models (LLMs) requires massive compute and full model access, which rules out state-of-the

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

MedPRESS: A Multi-turn Benchmark for Patient-Pressure-Induced Medical Sycophancy in LLMs

DGX agent

arXiv:2608.02520v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for health-related advice. Existing research measures their safety with static questions rather than

model-releasesarxiv-cs-cl
4 Aug 2026
Agents

MedTextWeaver: Procedural Knowledge Evolution in Agentic Medical Text Editing

DGX agent

arXiv:2602.00740v2 Announce Type: replace Abstract: Medical text editing is essential for improving communication among diverse stakeholders in clinical settings. However, adapting LLM agents to this

agentsarxiv-cs-cl
4 Aug 2026
Safety

MedUPS: Towards Diagnostic Assistance in Uncommon Medical Cases with Large Language Models

DGX agent

arXiv:2608.01012v1 Announce Type: new Abstract: Uncommon and off-guideline cases are difficult for clinical decision support, because physicians must make a series of management decisions under diagno

safetyarxiv-cs-cl
4 Aug 2026
Agents

MemoryForge: Synthesize Lifelong Memory for Human-Like LLM Agents

DGX agent

arXiv:2608.00007v1 Announce Type: new Abstract: Equipping Large Language Models (LLMs) with human-like personas is crucial for agentic applications, such as role-play and user simulation. Traditional

agentsarxiv-cs-cl
4 Aug 2026
Research

MemSIF: From Structured Interactions to Dual-Track Fact Memory for LLM Agents

DGX agent

arXiv:2608.01742v1 Announce Type: cross Abstract: Long-term memory is critical for LLM agents operating over long-horizon interactions. However, several persistent limitations of existing memory syste

researcharxiv-cs-cl
4 Aug 2026
Safety

Mind the Gap: Zero-Query Jailbreaks via Filter-Generator Discrepancy in Text-to-Image Systems

DGX agent

arXiv:2608.00973v1 Announce Type: new Abstract: Text-to-image (T2I) systems typically have prompt-level safety filters before the generator to block unsafe requests, yet such systems remain vulnerable

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

Morphology Aware Reversible Semantic Tokenization and Hierarchical Word Composition for Tamil Language Models

DGX agent

arXiv:2608.01153v1 Announce Type: new Abstract: Statistical subword tokenizers can process arbitrary text, but their units need not align with lexical or grammatical structure. This is especially impo

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

Native Multilingual Chain-of-Thought Reasoning in Low-Resource Southeast Asian Languages

DGX agent

arXiv:2608.00533v1 Announce Type: new Abstract: Large Language Models have achieved substantial progress in reasoning capabilities. Yet in low-resource native settings, many suffer from cross-lingual

safetyarxiv-cs-cl
4 Aug 2026
Research

Neural Circuit Function Inference with LLMs

DGX agent

arXiv:2608.00059v1 Announce Type: new Abstract: The success of connectome mapping now shifts the challenge of understanding the nervous system to the interpretation of neural circuits. Here, we devise

researcharxiv-cs-cl
4 Aug 2026
Safety

No One Wins in Nuclear War: A Social Simulation of Military Decision-making

DGX agent

arXiv:2608.01868v1 Announce Type: cross Abstract: WOPR is a social-simulation environment for studying how organizations make high-stakes decisions, built on a deterministic, replay-validated rules en

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

Not the Dimension, the Norm: What Matters in Gradient-Free Weight Perturbation of Language Models

DGX agent

arXiv:2608.01624v1 Announce Type: new Abstract: Adapting a language model to a task no longer requires training all of its weights, and a line of parameter-efficient methods has driven the trainable c

model-releasesarxiv-cs-cl
4 Aug 2026
Research

Observatorio Lazaro: A self-populating database of anglicism usage in the Spanish press

DGX agent

arXiv:2608.00713v1 Announce Type: new Abstract: This paper describes Observatorio Lazaro, a language resource that monitors unassimilated lexical borrowings (predominantly English lexical borrowings o

researcharxiv-cs-cl
4 Aug 2026
Model Releases

Obshazard-bench: Benchmarking Multimodal Foundation Models for Real-Time Disaster Intelligence from Raw Earth Observation Streams

DGX agent

arXiv:2608.00012v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are increasingly used to interpret Earth observation data, yet their capability to support real-world disaster

model-releasesarxiv-cs-cl
4 Aug 2026
Research

On the Wings of Imagination: Conflicting Script-based Multi-role Framework for Humor Caption Generation

DGX agent

arXiv:2602.06423v2 Announce Type: replace Abstract: Humor is a commonly used and intricate human language in daily life. Humor generation, especially in multi-modal scenarios, is a challenging task fo

researcharxiv-cs-cl
4 Aug 2026
Model Releases

OoO-Spec: Out-of-Order Semantic Speculation for Fast Tool Calling

DGX agent

arXiv:2608.00814v1 Announce Type: new Abstract: LLMs generate tool calls token by token, even though the function choice and argument values can often be predicted in parallel from the request and too

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution

DGX agent

arXiv:2608.00677v1 Announce Type: new Abstract: AI agents operate in persistent environments where early state changes can influence decisions far into the future. Unlike conventional language-model i

model-releasesarxiv-cs-cl
4 Aug 2026
Research

OpenDebateEvidence: A Massive-Scale Argument Mining and Summarization Dataset

DGX agent

arXiv:2406.14657v4 Announce Type: replace Abstract: We introduce OpenDebateEvidence, a comprehensive dataset for argument mining and summarization sourced from the American Competitive Debate communit

researcharxiv-cs-cl
4 Aug 2026
Model Releases

Opt.Gear Technical Report

DGX agent

arXiv:2608.01034v1 Announce Type: new Abstract: We introduce Opt.Gear, a foundation model designed for efficient on-device deployment, real-tim inference, and strong task capability. It includes a den

model-releasesarxiv-cs-cl
4 Aug 2026
Research

Orchestrating Dual-Boundaries: An Arithmetic Intensity Inspired Acceleration Framework for Diffusion Language Models

DGX agent

arXiv:2511.21759v2 Announce Type: replace Abstract: Diffusion-based large language models (dLLMs) have recently gained significant attention for their exceptional performance and inherent potential fo

researcharxiv-cs-cl
4 Aug 2026
Agents

OTAP: Structure-Aware Optimal Transport for Evaluating Planning and Execution in Agent Trajectories

DGX agent

arXiv:2607.17082v2 Announce Type: replace-cross Abstract: Large language model agents solve tasks by generating trajectories that interleave planning, tool calls, and intermediate results. Current eva

agentsarxiv-cs-cl
4 Aug 2026
Safety

PALMs: Using Multi Construct-Grounded Rationales for Modeling Population Preferences in LLMs

DGX agent

arXiv:2608.01458v1 Announce Type: new Abstract: Large language models are being extensively used to simulate individual user behavior, yet faithfully representing a population requires capturing the s

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

Passing Coarse Marginal Checks Can Be Cheap: Persona Mixtures and Imprecise Treatment-Response Estimates in an LLM Persona Panel

DGX agent

arXiv:2608.00979v1 Announce Type: cross Abstract: Large language models are increasingly used as synthetic research participants and are often validated by whether their marginal responses resemble hu

model-releasesarxiv-cs-cl
4 Aug 2026
Agents

PGMem: Tightly Coupled Persona-Memory Graph for Lifelong Personalized Agents

DGX agent

arXiv:2608.01708v1 Announce Type: new Abstract: Long-term personalized dialogue agents must track user preferences as their personas evolve. Existing memory systems organize past events well, but stor

agentsarxiv-cs-cl
4 Aug 2026
Research

PICTURE: Enhancing Theory-of-Mind in Large Language Models by Revealing, Not Hiding, Characters' Lack of Knowledge

DGX agent

arXiv:2608.01598v1 Announce Type: new Abstract: Simulating human-like Theory of Mind (ToM) has been a longstanding problem in natural language processing (NLP). To address this, existing works introdu

researcharxiv-cs-cl
4 Aug 2026
Research

PlainMedScale: A Corpus of Multi-Level Simplified Medical Texts in German and English

DGX agent

arXiv:2608.01158v1 Announce Type: new Abstract: We introduce PlainMedScale, a topic-aligned medical corpus spanning four levels of comprehensibility in German and English, drawn from MSD (professional

researcharxiv-cs-cl
4 Aug 2026
← Previous
1…1112131415…160
Next →