AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

A Shared Subcircuit Lets LLMs Count Down Across Tasks

DGX agent

arXiv:2607.12279v1 Announce Type: new Abstract: Writing a sentence of exactly twelve words; ending a DNA sequence at the right codon; formatting an ASCII table. These are all tasks that language model

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Agentic systems for breast cancer treatment recommendations

DGX agent

arXiv:2607.12051v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being explored for clinical decision support, but their reliability in complex oncology treatment planning

model-releasesarxiv-cs-cl
15 Jul 2026
Agents

Amplitude-Only FFN Intervention for Tool-Structured LLM Inference Method: Gated Evaluation Protocol, and Cross-Model Empirical Results

DGX agent

arXiv:2607.11183v2 Announce Type: replace Abstract: Large language models increasingly operate as tool-using agents, where small format, argument, or function-call errors can invalidate otherwise plau

agentsarxiv-cs-cl
15 Jul 2026
Model Releases

Belief-reality separation lives in routing over a shared value slot in language models

DGX agent

arXiv:2607.11945v1 Announce Type: new Abstract: Capable language models hold what a character believes apart from what is true: told 'Anna believes the cup is blue; in reality it is red,' they answer

model-releasesarxiv-cs-cl
15 Jul 2026
Research

Beyond Binary Detection: A Multi-Dimensional Taxonomy of Cancer Misinformation on Reddit

DGX agent

arXiv:2607.12383v1 Announce Type: new Abstract: Cancer-related discussions on social media provide an important space for information exchange and peer support, but also facilitate the spread of misin

researcharxiv-cs-cl
15 Jul 2026
Safety

Beyond Parallel Tracking: Interactive Multi-Feature Fusion Drives Semantic Reconstruction from Non-invasive Brain Recordings

DGX agent

arXiv:2607.12071v1 Announce Type: new Abstract: Continuous semantic reconstruction from non-invasive neural recordings remains limited by the representational mismatch between semantic feature spaces

safetyarxiv-cs-cl
15 Jul 2026
Tutorials

Can a Language Model Learn Facts Continually in Its Weights?

DGX agent

arXiv:2607.11020v2 Announce Type: replace Abstract: Continual learning promises a language model that keeps acquiring knowledge after training, with each new fact written into its weights. Whether wei

tutorialsarxiv-cs-cl
15 Jul 2026
Safety

Can LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment Reproduction

DGX agent

arXiv:2607.12835v1 Announce Type: new Abstract: Rubric-based evaluation is a promising approach for assessing open-ended outputs from LLM-based research agents, particularly in paper reproduction, whe

safetyarxiv-cs-cl
15 Jul 2026
Safety

CityBehavEx: A Scalable and Empirically Validated LLM-Assisted Urban Simulation Platform

DGX agent

arXiv:2607.12086v1 Announce Type: new Abstract: Recent LLM-based multi-agent urban simulators can generate semantically rich city routines, but they remain costly to scale and are often weakly validat

safetyarxiv-cs-cl
15 Jul 2026
Research

Entropy in Semantic Memory Navigation in Blind and Sighted Individuals: The Effect of Visual Experience

DGX agent

arXiv:2607.12185v1 Announce Type: new Abstract: Embodied accounts of semantic memory highlight the role of sensorimotor systems in acquiring and storing knowledge. Congenitally blind populations offer

researcharxiv-cs-cl
15 Jul 2026
Research

Entropy-Preserving Supervised Fine-Tuning via Adaptive Self-Distillation for Large Reasoning Models

DGX agent

arXiv:2602.02244v3 Announce Type: replace-cross Abstract: The standard post-training recipe for large reasoning models, supervised fine-tuning followed by reinforcement learning (SFT-then-RL), may lim

researcharxiv-cs-cl
15 Jul 2026
Model Releases

Epistemic Stance Flexibility Probing: Measuring Prompt-Conditioned Register Shift in Large Language Models

DGX agent

arXiv:2607.12739v1 Announce Type: new Abstract: A language model may be asked either what experts believe about a contested claim or what it believes about the claim itself. A trustworthy conversation

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

Evaluating Large Language Models on Misconceptions in Multi-Turn Medical Conversations

DGX agent

arXiv:2607.12884v1 Announce Type: new Abstract: Patients seeking medical information often ask questions that embed incorrect assumptions or misconceptions. In such cases, safe medical communication r

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

Extractable Memorization From First Principles

DGX agent

arXiv:2607.12649v1 Announce Type: cross Abstract: Recent work on extractable memorization in LLMs suffers from two contrasting validity problems. Some studies overstate extraction, e.g., relying on se

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

FairCoder: Probing LLM Bias in High-Stakes Decision Making via Coding Tasks

DGX agent

arXiv:2501.05396v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used in high-stakes decisions such as hiring and college admissions, making their social bias a critic

model-releasesarxiv-cs-cl
15 Jul 2026
Agents

Fine-Tuned Multi-Agent Framework for Detecting OCEAN in Life Narratives

DGX agent

arXiv:2607.12215v1 Announce Type: new Abstract: Accurately assessing personality from text is challenging because traits are latent, context-dependent, and often subtly expressed across long narrative

agentsarxiv-cs-cl
15 Jul 2026
Model Releases

FinResearchBench II: A Deep Research Benchmark with Consensus-Derived Gold Rubrics for Distinguishing Financial Report Quality

DGX agent

arXiv:2607.12252v1 Announce Type: new Abstract: Deep research agents are increasingly used to produce long-form financial reports, yet large-scale evaluation remains bottlenecked by the need for human

model-releasesarxiv-cs-cl
15 Jul 2026
Safety

From Sentiment to Actionable Insights: Public Sentiment Analysis of Advanced Air Mobility

DGX agent

arXiv:2606.20751v2 Announce Type: replace Abstract: Advanced Air Mobility (AAM) is an emerging low-altitude transportation system whose successful deployment depends on both technological progress and

safetyarxiv-cs-cl
15 Jul 2026
Research

From Words to Widgets for Controllable LLM Generation

DGX agent

arXiv:2604.10925v2 Announce Type: cross Abstract: Natural language remains the predominant way people interact with large language models (LLMs). However, users often struggle to precisely express and

researcharxiv-cs-cl
15 Jul 2026
Safety

Growing a Tail: Increasing Output Diversity in Large Language Models

DGX agent

arXiv:2411.02989v2 Announce Type: replace Abstract: How diverse are the outputs of large language models when diversity is desired? We examine the diversity of responses of several language models to

safetyarxiv-cs-cl
15 Jul 2026
Research

Hierarchical Latent Structures in Data Generation Process Unify Mechanistic Phenomena across Scale

DGX agent

arXiv:2603.06592v2 Announce Type: replace Abstract: Contemporary studies in mechanistic interpretability have uncovered many puzzling phenomena in the neural information processing of Transformer-base

researcharxiv-cs-cl
15 Jul 2026
Research

Hybrid Continual Learning for Low-Resource Australian Aboriginal Language Identification

DGX agent

arXiv:2607.11946v1 Announce Type: new Abstract: Language identification is an important step toward integrating endangered Australian Aboriginal languages (AALs) into speech technologies supporting la

researcharxiv-cs-cl
15 Jul 2026
Model Releases

KnowAct-GUIClaw: Know Deeply, Act Perfectly, Personal GUI Assistant with Self-Evolving Memory and Skill

DGX agent

arXiv:2607.12625v1 Announce Type: new Abstract: OpenClaw has emerged as a leading agent framework for complex task automation, yet it faces insufficient cross-platform GUI interaction support and a we

model-releasesarxiv-cs-cl
15 Jul 2026
Research

Knowledgeless Language Models: Suppressing Parametric Recall for Evidence-Grounded Language Modeling

DGX agent

arXiv:2607.12831v1 Announce Type: new Abstract: Language models encode substantial factual knowledge in their parameters, which can lead to unreliable behavior when this knowledge is outdated, incompl

researcharxiv-cs-cl
15 Jul 2026
Tutorials

Language Identification with Succinct Machine-Independent Traces

DGX agent

arXiv:2607.12443v1 Announce Type: new Abstract: Motivated by the power of large language models, there has been renewed interest in the Gold-Angluin model of language identification in the limit, with

tutorialsarxiv-cs-cl
15 Jul 2026
Model Releases

Learning Mechanistic Reasoning for Chemical Reactions with Large Language Models

DGX agent

arXiv:2607.12771v1 Announce Type: cross Abstract: Reaction mechanisms consist of the step-by-step sequences of elementary reactions that explain chemical transformations. Learning the mechanism logic

model-releasesarxiv-cs-cl
15 Jul 2026
Research

LLM Judges Can Be Too Generous When There Is No Reference Answer

DGX agent

arXiv:2607.12885v1 Announce Type: new Abstract: LLM judges are increasingly being used to evaluate open-ended model responses, often in no-reference settings where a ground-truth answer is unavailable

researcharxiv-cs-cl
15 Jul 2026
Model Releases

Lost in the Maze: Overcoming Context Limitations in Long-Horizon Agentic Search

DGX agent

arXiv:2510.18939v2 Announce Type: replace Abstract: Long-horizon agentic search requires iteratively exploring the web over long trajectories and synthesizing information across many sources, enabling

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

MAGE: Understanding Stability-Performance Trade-offs in Multi-component Prompt Optimization

DGX agent

arXiv:2607.11944v1 Announce Type: new Abstract: How do different components of iterative prompt optimization interact, and what happens when they are combined? We investigate this through MAGE (Memory

model-releasesarxiv-cs-cl
15 Jul 2026
Safety

Policy-Conditioned Constrained Decoding for Column-Level Access Control in Text-to-SQL

DGX agent

arXiv:2607.12341v1 Announce Type: new Abstract: Text-to-SQL is increasingly deployed across trust boundaries between data providers and users. Such deployment must balance three competing requirements

safetyarxiv-cs-cl
15 Jul 2026
Model Releases

Predict the Retrieval! Test time adaptation for Retrieval Augmented Generation

DGX agent

arXiv:2601.11443v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) has emerged as a powerful approach for enhancing large language models' question-answering capabilities through

model-releasesarxiv-cs-cl
15 Jul 2026
Research

QUBO-Optimized Evidence Selection for Retrieval-Augmented Question Answering with Unconventional Solvers

DGX agent

arXiv:2607.12334v1 Announce Type: new Abstract: Retrieval-augmented question answering depends on selecting evidence passages that jointly support answer generation. However, many RAG pipelines rely o

researcharxiv-cs-cl
15 Jul 2026
Applications

Rethinking Evaluation in Retrieval-Augmented Personalized Dialogue: A Cognitive and Linguistic Perspective

DGX agent

arXiv:2603.14217v3 Announce Type: replace Abstract: In cognitive science and linguistic theory, dialogue is not seen as a chain of independent utterances but rather as a joint activity sustained by co

applicationsarxiv-cs-cl
15 Jul 2026
Research

Ring-Zero: Scaling Zero RL to a Trillion Parameters for Emergent Reasoning

DGX agent

arXiv:2607.12395v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards without human-annotated data, often referred to as zero RL, has emerged as a powerful paradigm for elicit

researcharxiv-cs-cl
15 Jul 2026
Research

Segregate, Refine, Integrate: Decomposing Multimodal Fusion for Sentiment Analysis

DGX agent

arXiv:2607.12686v1 Announce Type: new Abstract: Multimodal fusion must simultaneously refine modality-specific signals and model cross-modal interactions; two competing objectives typically entangled

researcharxiv-cs-cl
15 Jul 2026
Agents

Speculate with Memory: Lossless Acceleration for LLM Agents

DGX agent

arXiv:2607.12236v1 Announce Type: cross Abstract: Speculative execution accelerates LLM agents by using a smaller, cheaper model to predict and pre-launch the next step while the environment is idle.

agentsarxiv-cs-cl
15 Jul 2026
Research

TAKE: Trajectory-Aware Knowledge Estimation for Text Dataset Distillation

DGX agent

arXiv:2607.11898v1 Announce Type: new Abstract: Large-scale text corpora have become a quiet bottleneck in modern NLP, not just in storage, but in the accumulated cost of training, fine-tuning, and co

researcharxiv-cs-cl
15 Jul 2026
Model Releases

The Capacity of Thought: Benchmarking Llama 3.2 in Semantic fMRI Neural Language Decoding and Improving the Huth Encoding-Model Baseline

DGX agent

arXiv:2607.12079v1 Announce Type: new Abstract: Decoding continuous language from fMRI signals remains a core challenge in non-invasive brain-computer interface research. We present two complementary

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

The Illusion of Robustness: Aggregate Accuracy Hides Prediction Flips under Task-Irrelevant Context

DGX agent

arXiv:2607.12963v1 Announce Type: new Abstract: As large language models (LLMs) grow more capable, they are increasingly deployed in context-rich settings where task inputs are often accompanied by lo

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

Token Reduction Is Not Cost Reduction

DGX agent

arXiv:2607.12161v1 Announce Type: new Abstract: Context-reduction layers for API-based coding agents, including command-output compressors, retrieval rankers, and payload-optimizing proxies, are usual

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

Transforming LLMs into Efficient Cross-Encoders via Knowledge Distillation for RAG Reranking

DGX agent

arXiv:2607.11933v1 Announce Type: new Abstract: Cross-encoders achieve high reranking accuracy in Retrieval-Augmented Generation (RAG) pipelines but impose quadratic inference costs that limit real-ti

model-releasesarxiv-cs-cl
15 Jul 2026
Research

Translation as a Computationally Efficient Bridge: Feasibility of English BERT for Low-Resource Languages

DGX agent

arXiv:2607.12612v1 Announce Type: new Abstract: BERT models have revolutionised Natural Language Processing (NLP) through their ability to process unstructured text across diverse domains. However, de

researcharxiv-cs-cl
15 Jul 2026
Safety

We Hebben Een Serieus Translatie: Modeling Intercomprehension as Probabilistic Inference

DGX agent

arXiv:2607.12169v1 Announce Type: new Abstract: Intercomprehension refers to partial intelligibility of an unfamiliar language (L2) by a speaker of a related language (L1). How is this zero-shot cross

safetyarxiv-cs-cl
15 Jul 2026
Model Releases

WikiSTAR: A System for Shedding Light on the Hidden History of Scientific Wikipedia Articles

DGX agent

arXiv:2607.12441v1 Announce Type: new Abstract: Wikipedia plays a key role in shaping public understanding of science, and its openly accessible revision history is a unique record of how scientific k

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

COALA: Robust Contextualized Speech-augmented Language Modeling for ASR via Contrastive Regularizer and Biasing Score Estimation

DGX agent

arXiv:2607.08117v1 Announce Type: new Abstract: Contextual biasing seeks to integrate external knowledge into automatic speech recognition (ASR) systems to accurately recognize domain-specific entitie

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

Cross-seed explainability using Procrustes-conditioned Joint End-to-end Top-K Sparse Autoencoders

DGX agent

arXiv:2607.08499v1 Announce Type: new Abstract: We present a Procrustes-conditioned Joint End-to-end Top-K Sparse Autoencoder (SAE) for extracting cross-seed universal features from independently trai

model-releasesarxiv-cs-cl
10 Jul 2026
Agents

DeepSearch-World: Self-Distillation for Deep Search Agents in a Verifiable Environment

DGX agent

arXiv:2607.07820v1 Announce Type: new Abstract: Training tool-use agents to improve from their own experience remains challenging, as supervised fine-tuning relies on fixed teacher-distilled trajector

agentsarxiv-cs-cl
10 Jul 2026
Safety

Detecting Ladder Logic Bombs in IEC 61131-3 PLC Programs using ESBMC-PLC+: A Formal Verification Approach with Trigger Synthesis

DGX agent

arXiv:2607.08417v1 Announce Type: new Abstract: A Ladder Logic Bomb (LLB) is malicious control logic in a Programmable Logic Controller (PLC) program that lies dormant until a trigger activates a payl

safetyarxiv-cs-cl
10 Jul 2026
← Previous
1…2728293031…161
Next →