AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design

DGX agent

arXiv:2603.01229v3 Announce Type: replace Abstract: Robotic manipulation policies have made rapid progress in recent years, yet most existing approaches give limited consideration to memory capabiliti

model-releasesarxiv-cs-ro
31 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Safety-Gated Agentic Supervisory Control on a Coupled Distillation Benchmark: Regime Map, Auditable Gate, and Co-Design Findings

DGX agent

arXiv:2607.27849v1 Announce Type: cross Abstract: An open-weight LLM can write composition setpoints every five minutes. What a plant still needs is a hard check: named constraints, logged margins, an

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Same Facts, Different Diagnosis: Measuring and Mitigating Narrative Anchoring in Clinical Language Models

DGX agent

arXiv:2607.27384v1 Announce Type: new Abstract: Large language models used for clinical diagnostic reasoning are sensitive to sociolinguistic register, not just clinical content. We term this failure

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

SARC-DQ: Runtime Data-Quality Gating for Agentic AI: Silent Evidence Defects, the Incompetence Shield, and Downstream-Only Remediation

DGX agent

arXiv:2607.26313v1 Announce Type: cross Abstract: Agentic systems act, so a defect in the evidence they retrieve becomes a wrong action with a currency cost. The most dangerous enterprise defects are

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Scaling medical imaging report generation with multimodal reinforcement learning

DGX agent

arXiv:2601.17151v2 Announce Type: replace-cross Abstract: Frontier models have demonstrated remarkable capabilities in understanding and reasoning with natural-language text, but they still exhibit ma

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Scaling Vision-Language Models Is Not Enough to Mitigate Bias

DGX agent

arXiv:2607.28211v1 Announce Type: new Abstract: Vision-Language Models (VLMs) such as CLIP are now foundational to multimodal systems, yet their robustness to spurious correlations remains poorly unde

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Select or Project? Evaluating Lower-dimensional Vectors for LLM Training Data Explanations

DGX agent

arXiv:2601.16651v3 Announce Type: replace Abstract: Gradient-based methods for instance-based explanation for large language models (LLMs) are hindered by the immense dimensionality of model gradients

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Selecting Open-Weight Language Models for Zero-Shot Intent Classification: A Systematic Evaluation of 41 Models

DGX agent

arXiv:2607.27421v1 Announce Type: new Abstract: Intent classification is a core component of task-oriented dialogue systems, yet practitioners have limited systematic guidance for selecting deployable

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

SemAnCorr: Semantic Anchored Correspondence for Zero-Shot Manipulation Skill Transfer

DGX agent

arXiv:2607.28382v1 Announce Type: new Abstract: Transferring manipulation skills across object instances that share functionality but differ in geometry remains a fundamental challenge in robot learni

model-releasesarxiv-cs-ro
31 Jul 2026
Model Releases

Shared Symbolic Backbones for Physically Consistent Multi-Output Symbolic Regression

DGX agent

arXiv:2607.26528v1 Announce Type: cross Abstract: Symbolic regression provides analytical expressions, but it is usually applied one output at a time. This is limiting in process systems, where state

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Sign Language Question Answering: A New Task, Benchmark, and Baseline for Sign Language Understanding

DGX agent

arXiv:2607.27826v1 Announce Type: cross Abstract: Recent advances in sign language (SL) understanding (SLU) have led to remarkable progress in tasks such as continuous SL recognition and SL translatio

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Simplifying Neural Networks During Training

DGX agent

arXiv:2607.27854v1 Announce Type: cross Abstract: Understanding and exploiting the training dynamics of overparameterized deep neural networks remains a central challenge in modern machine learning. R

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Space2Ground 2.0: A Multi-Source Dataset and Framework for Agricultural Monitoring through Fusion of Street-Level and Satellite Imagery

DGX agent

arXiv:2607.28247v1 Announce Type: new Abstract: Accurate and scalable parcel-level agricultural monitoring remains challenging because satellite Earth Observation alone provides only an overhead persp

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

StealthBench: Measuring Operational Stealth in Autonomous Offensive-Security Agents

DGX agent

arXiv:2607.26314v1 Announce Type: cross Abstract: Stealth, the discipline of achieving an objective without revealing your presence, capabilities, or collected intelligence, is what separates sophisti

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

STEREODISCO: Discovering Stereotypicality in LLMs

DGX agent

arXiv:2607.27824v1 Announce Type: cross Abstract: LLMs encode, convey, and perpetuate stereotypes. Prior computational research focuses on a small set of semantic axes investigated in social psycholog

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Subtract or Replay? Exact Deletion from Language-Model Memory

DGX agent

arXiv:2607.27539v1 Announce Type: cross Abstract: Exact deletion from persistent language-model memory depends on how that memory represents a record. Addressable influence can be removed by algebraic

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Sympathetic Framing: Evaluating AI Alignment across Sociodemographic Groups

DGX agent

arXiv:2607.27232v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly shaping how we consume information and form our worldview. This raises concerns beyond bias in AI: do LLMs

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

TEA-AgriVLN: Traversability Estimation Alarm for Agricultural Vision-and-Language Navigation

DGX agent

arXiv:2607.28474v1 Announce Type: new Abstract: Vision-and-Language Navigation in Continuous Environments (VLN-CE) requires an agent to follow a natural language instruction, predicting a sequence of

model-releasesarxiv-cs-ro
31 Jul 2026
Model Releases

The Convergence Behavior of Adam under Heavy-Tailed Noise

DGX agent

arXiv:2607.27383v1 Announce Type: new Abstract: We establish the first convergence guarantees for the plain vector-form Adam optimizer under heavy-tailed stochastic noise. While several Adam variants

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

THGFM: Dual-Branch Temporal Heterogeneous Graph Fusion Model

DGX agent

arXiv:2607.27303v1 Announce Type: cross Abstract: Temporal heterogeneous graphs offer a natural abstraction for dynamic relational systems in which diverse node and relation types co-exist and evolve

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Think with Extra-Image: A Farmland Segmentation Agent Driven by Spatio-Temporal Information Gain

DGX agent

arXiv:2607.28186v1 Announce Type: new Abstract: Existing farmland remote sensing image (FRSI) segmentation follows a 'Think with Intra-Image' paradigm, assuming that the current image contains suffici

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Thinking Once Is Enough: Intermediate-Layer Evidence Routing for High-Resolution VQA

DGX agent

arXiv:2607.27830v1 Announce Type: new Abstract: High-resolution visual question answering (HR-VQA) is often treated as a problem of insufficient evidence acquisition, where failing multimodal large la

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Tight Sample Complexity for Low-Rank Adaptation: Matching Bounds and Rank Selection

DGX agent

arXiv:2607.27680v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) has become the standard mechanism for fine-tuning large pretrained models, yet its statistical properties remain only parti

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Towards Generalized Synapse Detection Across Invertebrate Species

DGX agent

arXiv:2509.17041v2 Announce Type: replace Abstract: Behavioural differences across organisms, whether healthy or pathological, are closely tied to the structure of their neural circuits. Yet, the fine

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Towards Stability of Parameter-Free Optimization

DGX agent

arXiv:2405.04376v4 Announce Type: replace Abstract: Hyperparameter tuning, particularly the selection of an appropriate learning rate in adaptive gradient training methods, remains a challenge. To add

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Towards Unified Multimodal Misinformation Detection in Social Media: A Benchmark Dataset and Baseline

DGX agent

arXiv:2509.25991v3 Announce Type: replace-cross Abstract: Detecting deceptive multimodal content on social media has become an increasingly important problem. Two major types of deception dominate: hu

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

TraceCoder: Explainable and Auditable Code Generation with Position-Key Snippet Versioning

DGX agent

arXiv:2607.26307v1 Announce Type: new Abstract: Contemporary LLM-based coding agents produce code as black-box outputs: the rationale behind each line is hidden, the evolution of the code through benc

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

TriShield: Zero-Utility-Loss Defense Against Privacy Backdoors in Federated Language Model Fine-Tuning via Orthogonal Gradient Projection and Optimizer State Entanglement

DGX agent

arXiv:2607.27940v1 Announce Type: cross Abstract: Federated fine-tuning of large language models (LLMs) enables collaborative training without exposing raw data. However, a recent attack, NeuroImprint

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Tycho: Active Abstraction with Programmatic World Models for ARC-AGI-3

DGX agent

arXiv:2607.28287v1 Announce Type: cross Abstract: ARC-AGI-3 turns abstraction into an interactive problem of skill acquisition. A player must infer an unfamiliar game's rules, hidden state, and goal w

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

UrbanDS: A Graph-Guided LLM Multi-Agent System for Data-Intensive Urban Tasks

DGX agent

arXiv:2607.26724v1 Announce Type: new Abstract: Large language model (LLM) agents have been widely applied in automating data science tasks. However, existing methods typically rely on a limited set o

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Using Large Language Models for Idea Generation in Innovation

DGX agent

arXiv:2607.27553v1 Announce Type: cross Abstract: This research evaluates the efficacy of large language models (LLMs) in generating new product ideas. To do so, we compare three pools of ideas for ne

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

VESTIGE: A Knowledge-Guided Masking Strategy for Corruption-Aware Fine-Tuning of Genomic Transformers, Validated on Ancient DNA Reconstruction

DGX agent

arXiv:2607.27712v1 Announce Type: new Abstract: Standard masked-language-model fine-tuning applies a uniform masking probability across every token position, assuming reconstruction difficulty is posi

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

What Does It Take to Detect an AI Agent? Minimal Feature Sets for Behavioral Detection under Browser Automation

DGX agent

arXiv:2607.26935v1 Announce Type: new Abstract: Bot detectors deployed at scale treat traffic as binary: human or bot. This assumption breaks when AI agents browse the web through browser automation,

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

What Makes Deep Learning Work for Traditional Chinese Medicine Tongue Diagnosis? A Comprehensive Ablation Study

DGX agent

arXiv:2607.28148v1 Announce Type: new Abstract: Deep learning has shown promise for automated tongue diagnosis in traditional Chinese medicine (TCM), yet the design space remains underexplored. We con

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

When Does Muon Help Agentic Reinforcement Learning?

DGX agent

arXiv:2607.16169v3 Announce Type: replace Abstract: Muon is competitive with AdamW in large-scale pre-training, but its operating regime in reinforcement-learning post-training remains unclear. We map

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

WhisperRec: Latent Reasoning for Efficient Foundation Recommendation Models

DGX agent

arXiv:2607.26621v2 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong reasoning capabilities, motivating their adoption as backbones for foundation recommendation mod

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Why Are GUI Agents Correct but Late? Decode on the Decision-Time Critical Path, Tested with Pre-Compiled Policy Trees

DGX agent

arXiv:2607.28399v1 Announce Type: new Abstract: Computer-use agents often fail on transient GUI events because they produce the correct action only after the relevant window has already closed. We ide

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Would You Walk to the Car Wash? Revealing the Salience Bias of Large Language Models in Commonsense Reasoning

DGX agent

arXiv:2607.28478v1 Announce Type: new Abstract: As large language models (LLMs) continue to advance in complex reasoning tasks, they have learned to heavily prioritize explicit conditions provided in

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

ZUNA1.1: A more flexible EEG foundation model for Denoising and Super-resolution

DGX agent

arXiv:2607.27308v1 Announce Type: new Abstract: We introduce ZUNA1.1, a 380M-parameter diffusion autoencoder for flexible EEG signal reconstruction. ZUNA1.1 is capable of reconstructing variable lengt

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

AdaMARP: An Adaptive Multi-Agent Interaction Framework for General Immersive Role-Playing

DGX agent

arXiv:2601.11007v2 Announce Type: replace-cross Abstract: LLM role-playing aims to portray arbitrary characters in interactive narratives, yet existing systems often suffer from limited immersion and

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Aligning LLM-Simulated and Human Examinees for Psychometric Calibration: A Cognitive Diagnostic Profiling Approach

DGX agent

arXiv:2607.26317v1 Announce Type: cross Abstract: Psychometric calibration for educational tests typically requires costly human response data. Large language models (LLMs) simulated examinees offer a

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Amortized Moment Matching for Visual Generation

DGX agent

arXiv:2607.26860v1 Announce Type: new Abstract: We propose amortized moment matching, utilizing neural networks to learn data moments as distributional training signals. By casting diffusion denoisers

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

APEX-Accounting

DGX agent

arXiv:2607.27189v1 Announce Type: new Abstract: We introduce APEX-Accounting, a benchmark built by Mercor in partnership with Ramp, to assess whether frontier models can do the real work of accountant

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

ARC-Encoder: learning compressed text representations for large language models

DGX agent

arXiv:2510.20535v2 Announce Type: replace Abstract: Recent techniques such as retrieval-augmented generation or chain-of-thought reasoning have led to longer contexts and increased inference costs. Co

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Archetypes or ability? Clustering for modelling student mathematical competence

DGX agent

arXiv:2607.26063v1 Announce Type: cross Abstract: Personalised learning systems often assume that mathematical ability is combined of discrete abilities, acquired sequentially and dependent upon first

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Automorphism-Induced Non-Canonicity in Top-k Explanations of Graph Neural Networks

DGX agent

arXiv:2607.26344v1 Announce Type: new Abstract: A gradient-based GNN explainer given a molecule with two chemically equivalent nitro groups assigns them attribution scores that are equal to the last b

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

BayesAME: Bayesian Active Model Evaluation

DGX agent

arXiv:2607.27023v1 Announce Type: new Abstract: Evaluating large generative models across benchmarks is time-consuming and computationally expensive. This drives the need for methods that can estimate

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Between Gradient and Natural Gradient: A Continuum of LoRA Initializations

DGX agent

arXiv:2607.26247v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) fine-tunes large pretrained models at a fraction of the cost of full fine-tuning, but its performance depends strongly on how

model-releasesarxiv-cs-lg
30 Jul 2026
← Previous
1…4243444546…357
Next →