AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Safety

MedGuards: Multi-Agent System for Reliable Medical Error Detection and Correction

DGX agent

arXiv:2606.25651v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed in healthcare settings, accurate error detection and correction in generated or existing text

safetyarxiv-cs-cl
25 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

MedLayBench-V: A Large-Scale Benchmark for Expert-Lay Semantic Alignment in Medical Vision Language Models

DGX agent

arXiv:2604.05738v2 Announce Type: replace Abstract: Medical Vision-Language Models (Med-VLMs) have achieved expert-level proficiency in interpreting diagnostic imaging. However, current models are pre

model-releasesarxiv-cs-cl
25 Jun 2026
Agents

Membox: Weaving Topic Continuity into Long-Range Memory for LLM Agents

DGX agent

arXiv:2601.03785v3 Announce Type: replace Abstract: Long-term human-agent dialogues are organized by topic continuity: adjacent turns often develop the same goal, plan, problem, or event, while relate

agentsarxiv-cs-cl
25 Jun 2026
Agents

Memory Makes the Difference: Evaluating How Different Memory Roles Shape Conversational Agents

DGX agent

arXiv:2606.25361v1 Announce Type: new Abstract: Prior research on memory mechanism in RAG-based conversational system has emphasized how memory is stored and retrieved. However, far less is known abou

agentsarxiv-cs-cl
25 Jun 2026
Model Releases

Multilingual Hematology Visual Question Answering Dataset

DGX agent

arXiv:2606.25246v1 Announce Type: cross Abstract: Vision Language Models (VLMs) have shown promising capabilities in medical image analysis by jointly understanding visual and textual information for

model-releasesarxiv-cs-cl
25 Jun 2026
Safety

Narrative Feature or Structured Feature? A Study of Large Language Models to Identify Cancer Patients at Risk of Heart Failure

DGX agent

arXiv:2403.11425v4 Announce Type: replace-cross Abstract: Cancer treatments are known to introduce cardiotoxicity, negatively impacting outcomes and survivorship. Identifying cancer patients at risk o

safetyarxiv-cs-cl
25 Jun 2026
Model Releases

Natural Ungrokking: Asymmetric Control of Which Rules Survive Pretraining

DGX agent

arXiv:2606.26050v1 Announce Type: cross Abstract: Midway through an ordinary pretraining run, a small language model learns the pronoun-gender rule: cued with a girl's name ('Sue cried because'), it r

model-releasesarxiv-cs-cl
25 Jun 2026
Safety

Neural Machine Translation for Low-Resource Tangkhul--English

DGX agent

arXiv:2606.25365v1 Announce Type: new Abstract: We present a study on low-resource machine translation for the Tangkhul-English (nmf-en) language pair. Tangkhul is a severely under-resourced Tibeto-Bu

safetyarxiv-cs-cl
25 Jun 2026
Research

Neural Scaling Universality: If Exponents Are Fixed, Time to Understand Coefficients

DGX agent

arXiv:2606.25008v1 Announce Type: cross Abstract: Neural scaling laws describe how pre-training loss decays as power laws with training time, model size, and compute. This position paper argues that t

researcharxiv-cs-cl
25 Jun 2026
Safety

OPERA: Aligning Open-Ended Reasoning via Objective Perplexity-based Reinforcement Learning

DGX agent

arXiv:2606.25757v1 Announce Type: new Abstract: Reinforcement Learning (RL) has enabled LLMs to excel in objective reasoning tasks such as mathematics and code generation. However, applying RL to open

safetyarxiv-cs-cl
25 Jun 2026
Research

Optimizing Abstractive Summarization With Fine-Tuned PEGASUS

DGX agent

arXiv:2606.25462v1 Announce Type: new Abstract: Abstractive text summarization is the technique of generating a short and concise summary comprising the salient ideas of a source text without making a

researcharxiv-cs-cl
25 Jun 2026
Research

Overview of HIPE-2026: Person-Place Relation Extraction from Multilingual Historical Texts

DGX agent

arXiv:2606.25935v1 Announce Type: new Abstract: Was this person ever at that place, and if so, when? Answering such questions from noisy, multilingual historical documents is the central challenge of

researcharxiv-cs-cl
25 Jun 2026
Safety

Paid Voices vs. Public Feeds: Interpretable Cross-Platform Theme-Based Analysis of Climate Discourse

DGX agent

arXiv:2601.13317v2 Announce Type: replace Abstract: Climate discourse online shapes public understanding of climate change and informs political and policy debate, yet it unfolds across structurally d

safetyarxiv-cs-cl
25 Jun 2026
Model Releases

Perfect Detection, Failed Control: The Geometry of Knowing vs. Steering in Language Models

DGX agent

arXiv:2606.24952v1 Announce Type: new Abstract: A central aspiration of mechanistic interpretability is controllability: if we know where a behavior is represented in a model's activations, we should

model-releasesarxiv-cs-cl
25 Jun 2026
Agents

PhoneBuddy: Training Open Models for Agentic Phone Use

DGX agent

arXiv:2606.23049v2 Announce Type: replace Abstract: Phones are becoming an important execution surface for general-purpose agents, but training open models for reliable phone use remains difficult bec

agentsarxiv-cs-cl
25 Jun 2026
Model Releases

PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models

DGX agent

arXiv:2606.25442v1 Announce Type: new Abstract: Safety alignment of large language models (LLMs) typically depends on high-quality supervision data, such as safe demonstrations or preference pairs. Ho

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Position: Reasoning After Perception Means Reasoning Without Vision

DGX agent

arXiv:2507.16863v2 Announce Type: replace-cross Abstract: A common belief in multimodal research is that the perceptual weaknesses of vision--language models can be compensated by stronger language re

researcharxiv-cs-cl
25 Jun 2026
Model Releases

Privacy-Aware Visual Language Models

DGX agent

arXiv:2405.17423v4 Announce Type: replace-cross Abstract: As Visual Language Models (VLMs) become increasingly embedded in everyday applications, ensuring they can recognise and appropriately handle p

model-releasesarxiv-cs-cl
25 Jun 2026
Applications

Probing in the Wild: A Case Study of Self-Supervised Speech Representations on Mandarin Sub-dialects with Unsupervised Articulatory Analysis

DGX agent

arXiv:2606.25459v1 Announce Type: new Abstract: While self-supervised speech models have achieved strong performance across speech tasks, relatively little is known about how their internal phonetic r

applicationsarxiv-cs-cl
25 Jun 2026
Model Releases

RAS: Measuring LLM Safety Through Refusal Alignment

DGX agent

arXiv:2606.25750v1 Announce Type: cross Abstract: Safety evaluation of large language models (LLMs) is commonly performed by querying models with unsafe or jailbreak prompts and judging whether their

model-releasesarxiv-cs-cl
25 Jun 2026
Agents

RAVEN: Long-Horizon Reasoning & Navigation with a Visuo-Spatio-Temporal Memory

DGX agent

arXiv:2606.25206v1 Announce Type: cross Abstract: Long-term robot deployment requires a compact and scalable memory that preserves fine-grained visual semantics, grounds observations in space and time

agentsarxiv-cs-cl
25 Jun 2026
Model Releases

Real-Time Voice AI Hears but Does Not Listen

DGX agent

arXiv:2606.26083v1 Announce Type: new Abstract: Speech conveys information through both words and vocal delivery. We evaluate four leading production realtime voice systems-OpenAI's GPT Realtime 2, Go

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Reclaim Evaluation: A Lossy Memory Is Worse Than an Empty One

DGX agent

arXiv:2606.25449v1 Announce Type: new Abstract: A language model's memory can be worse than having no memory at all. Give a model a memory that kept a wrong conclusion but dropped the work behind it,

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Reinforcement Learning Improves Traversal of Parametric Knowledge in LLMs

DGX agent

arXiv:2511.05933v2 Announce Type: replace Abstract: Reinforcement learning (RL) is often credited with improving language model reasoning at the expense of knowledge. We challenge this narrative by sh

researcharxiv-cs-cl
25 Jun 2026
Research

Riazi-8B: An Urdu Large Language Model for Mathematical Reasoning

DGX agent

arXiv:2606.25568v1 Announce Type: new Abstract: Recent LLMs demonstrate strong mathematical reasoning capabilities, but existing gains rely heavily on English-centric training resources and benchmarks

researcharxiv-cs-cl
25 Jun 2026
Model Releases

Robustness assessment of large audio language models in multiple-choice evaluation

DGX agent

arXiv:2510.04584v2 Announce Type: replace Abstract: Recent advances in large audio language models (LALMs) have primarily been assessed using a multiple-choice question answering (MCQA) framework. How

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Same Evidence, Different Answer: Auditing Order Sensitivity in Multimodal Large Language Models

DGX agent

arXiv:2606.26079v1 Announce Type: new Abstract: Standard benchmarks for multimodal large language models (MLLMs) score each item on one canonical ordering and miss whether order-irrelevant shuffling c

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

SARA: Unlocking Multilingual Knowledge in Mixture-of-Experts via Semantically Anchored Routing Alignment

DGX agent

arXiv:2606.25821v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) architectures have emerged as an increasingly influential paradigm as they offer a strategic balance between parameter s

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Sarashina2.2-TTS: Tackling Kanji Polyphony in Japanese Speech Generation via Data Scaling and Targeted Data Synthesis

DGX agent

arXiv:2606.25369v1 Announce Type: cross Abstract: While large language model (LLM)-based text-to-speech (TTS) systems have achieved high-quality speech synthesis, most existing systems focus on Englis

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Scale or Reason? A Compute-Equivalent Analysis of Reasoning Distillation

DGX agent

arXiv:2509.22193v2 Announce Type: replace Abstract: Distilling reasoning traces from strong teacher models has become the standard recipe for building capable small language models. Yet reasoning trac

researcharxiv-cs-cl
25 Jun 2026
Local Ai

Security and Privacy in Retrieval-Augmented Generation: Architectures, Threats, Defenses, and Future Directions for Building Trustworthy Systems

DGX agent

arXiv:2606.25533v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) has emerged as a dominant paradigm for enhancing large language models with external knowledge. By coupling retri

local-aiarxiv-cs-cl
25 Jun 2026
Model Releases

SFL-MTSC: Leveraging Semantic Frame-Level Multi-Task Self-Consistency for Robust Multi-Intent Spoken Language Understanding

DGX agent

arXiv:2606.25552v1 Announce Type: new Abstract: Prompt-based spoken language understanding (SLU) with large language models (LLMs) often suffers from inconsistent intent--slot structures due to decodi

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Small edits, large models: How Wikipedia advocacy shapes LLM values

DGX agent

arXiv:2606.24890v1 Announce Type: new Abstract: Can a small group of volunteers shape how AI systems discuss animal welfare, just by editing Wikipedia? We show that they can. Wikipedia appears in near

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Space-Efficient Language Generation in the Limit

DGX agent

arXiv:2606.25777v1 Announce Type: cross Abstract: We initiate a resource-aware theory of extit{language generation in the limit} under the minimal constraint of space efficiency. In our framework, a l

researcharxiv-cs-cl
25 Jun 2026
Research

Spam and Sentiment Detection in Arabic Tweets Using MARBERT Model

DGX agent

arXiv:2606.25495v1 Announce Type: new Abstract: Saudi Telecom Company (STC) is among the most popular companies in Saudi Arabia, with many customers. Yet, there is still a big room for improvement in

researcharxiv-cs-cl
25 Jun 2026
Model Releases

SPARC: Separating Perception And Reasoning Circuits for Test-time Scaling of VLMs

DGX agent

arXiv:2602.06566v3 Announce Type: replace-cross Abstract: Despite recent successes, test-time scaling -- i.e., dynamically expanding the token budget during inference as needed -- remains brittle for

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Speech Codec Probing from Semantic and Phonetic Perspectives

DGX agent

arXiv:2603.10371v2 Announce Type: replace-cross Abstract: Speech tokenizers are essential for connecting speech to large language models (LLMs) in multimodal systems. Speech tokenizers are expected to

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

SpeechEQ: Benchmarking Emotional Intelligence Quotient in Socially Aware Voice Conversational Models

DGX agent

arXiv:2606.25990v1 Announce Type: new Abstract: As multimodal conversational systems increasingly engage in spoken interaction, their ability to navigate paralinguistic social cues has become a critic

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Staying In Character: Perspective-Bounded Memory For Book-Based Role-Playing Agents

DGX agent

arXiv:2606.25632v1 Announce Type: new Abstract: Recent LLM role-playing systems build character agents from novels by extracting characters, scenes, and relations. Yet long-narrative role-playing suff

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Story Operators: Decomposing the Original o Sequel Transformation in Embedding Space

DGX agent

arXiv:2606.25379v1 Announce Type: new Abstract: I treat a book as a point in a sentence-embedding space and a literary transformation as an operation on points. Given an original novel and its sequel,

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Streaming-dLLM: Accelerating Diffusion LLMs via Suffix Pruning and Dynamic Decoding

DGX agent

arXiv:2601.17917v3 Announce Type: replace-cross Abstract: Diffusion Large Language Models (dLLMs) offer a compelling paradigm for natural language generation, leveraging parallel decoding and bidirect

researcharxiv-cs-cl
25 Jun 2026
Model Releases

SyncLoop: A Multimodal Dual-Loop Framework for Self-Improving Mathematical Reasoning

DGX agent

arXiv:2507.16518v3 Announce Type: replace-cross Abstract: Recent advances in multimodal large language models (MLLMs) have shown impressive reasoning capabilities. However, further enhancing existing

model-releasesarxiv-cs-cl
25 Jun 2026
Research

The cognitive, affective, and behavioral expression of self-stigma among people who use drugs in online substance use communities

DGX agent

arXiv:2606.25143v1 Announce Type: new Abstract: Objectives: To develop a codebook for self-stigma across cognitive, affective, and behavioral domains, and to estimate the prevalence, co-occurrence, an

researcharxiv-cs-cl
25 Jun 2026
Research

The Generalization Spectrum: A Chromatographic Approach to Evaluating Learning Algorithms

DGX agent

arXiv:2606.25450v1 Announce Type: cross Abstract: Traditional evaluations measure a learning algorithm's final performance on an i.i.d. test set, reducing learning to a single aggregate score. This ap

researcharxiv-cs-cl
25 Jun 2026
Safety

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems

DGX agent

arXiv:2606.24937v1 Announce Type: cross Abstract: The Hitchhiker's Guide to Agentic AI is a comprehensive practitioner's reference for building autonomous AI systems. The book covers the full stack fr

safetyarxiv-cs-cl
25 Jun 2026
Research

The Interplay of Harness Design and Post-Training in LLM Agents

DGX agent

arXiv:2606.25447v1 Announce Type: cross Abstract: Tool-integrated LLM agents are often wrapped within a harness: the scaffolding that determines which tools are exposed, how they are described, and wh

researcharxiv-cs-cl
25 Jun 2026
Safety

The Tatoxa System for Text Detoxification in Low-Resource Languages: The Case of Tatar

DGX agent

arXiv:2606.26015v1 Announce Type: new Abstract: Text detoxification, the automated detection and mitigation of abusive and harmful content, is essential for ensuring the safety of online communities a

safetyarxiv-cs-cl
25 Jun 2026
Model Releases

Three Buddhist Vocabularies: Computational Stylometry of the English Pali Canon across Sutta, Vinaya, and Abhidhamma

DGX agent

arXiv:2606.25372v1 Announce Type: new Abstract: We present a computational stylometric analysis of the Tipitaka across all three Pitakas in English translation, extending earlier work on the Sutta Pit

model-releasesarxiv-cs-cl
25 Jun 2026
← Previous
1…4344454647…161
Next →