AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Optimization-based Online Conformal Prediction for Multi-step Forecasting

DGX agent

arXiv:2508.13362v2 Announce Type: replace Abstract: Conformal prediction (CP) is well-suited for uncertainty quantification in time series forecasting due to its distribution-free coverage guarantees.

model-releasesarxiv-cs-lg
10 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Overcoming Rank Collapse in Feedback Alignment

DGX agent

arXiv:2606.11123v1 Announce Type: new Abstract: Backpropagation (BP) is widely viewed as biologically implausible, in part because it requires feedback weights to be the transpose of forward weights f

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

P3D-Bench: Benchmarking MLLMs for Parametric 3D Generation and Structural Reasoning

DGX agent

arXiv:2606.11152v1 Announce Type: new Abstract: Multimodal large language models can write code to produce complex programs as well as use programs to do 3D modeling, which opens up a new avenue for 3

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Parallel Causal Associative Fields: Gated Sparse Memory for Long-Context Language Modeling

DGX agent

arXiv:2606.10435v1 Announce Type: cross Abstract: Transformers achieve strong language modeling performance by providing direct token-to-token communication paths, but causal self-attention scales qua

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Parametric Knowledge is Not All You Need: Toward Honest Large Language Models via Retrieval of Pretraining Data

DGX agent

arXiv:2601.21218v2 Announce Type: replace Abstract: Large language models (LLMs) are highly capable of answering questions, but they are often unaware of their own knowledge boundary, i.e., knowing wh

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

PhantomBench: Benchmarking the Non-existential Threat of Language Models

DGX agent

arXiv:2606.11105v1 Announce Type: cross Abstract: Hallucinations, where language models (LMs) generate factually ungrounded responses, pose serious risks, as users tend to blindly rely on them. This i

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Piper: A Programmable Distributed Training System

DGX agent

arXiv:2606.11169v1 Announce Type: cross Abstract: Large-scale model training increasingly relies on composing multiple parallelism strategies, such as data, pipeline, and expert parallelism, together

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

POPSICLE: Benchmark Datasets for Segmentation and Localization in CryoET

DGX agent

arXiv:2606.10255v1 Announce Type: cross Abstract: Cryo-electron tomography (cryoET) has emerged as a powerful tool in structural and cellular biology by enabling direct visualization of macromolecular

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

PreAct-Bench: Benchmarking Predictive Monitoring in LLMs

DGX agent

arXiv:2606.09890v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents capable of executing multi-step action trajectories toward a given objecti

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

ProbeLLM: Automating Principled Diagnosis of LLM Failures

DGX agent

arXiv:2602.12966v2 Announce Type: replace Abstract: Understanding how and why large language models (LLMs) fail is becoming a central challenge as models rapidly evolve and static evaluations fall beh

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Quality Is Not a Safety Proxy Under Quantization

DGX agent

arXiv:2606.10154v1 Announce Type: new Abstract: Quantized checkpoints are often screened first with quality metrics and only later, if at all, with direct safety tests. This paper audits that shortcut

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Quantifying Uncertainty in AI Visibility: A Statistical Framework for Generative Search Measurement

DGX agent

arXiv:2603.08924v2 Announce Type: replace-cross Abstract: AI-powered answer engines are inherently non-deterministic: identical queries submitted at different times can produce different responses and

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Quo Vadis, Visual In-Context Learning? A Unified Benchmark Across Domains and Tasks

DGX agent

arXiv:2606.10967v1 Announce Type: new Abstract: Visual in-context learning has been proposed as a pathway towards dynamic models that can generate predictions based on a provided context and thereby c

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Rank Collapse, Fixed Points, and the Renormalization Group Structure of MLP Residual Networks

DGX agent

arXiv:2606.10324v1 Announce Type: new Abstract: The analogy between deep neural network forward passes and renormalization group (RG) flows has been repeatedly noted in the literature, but existing tr

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

RAT: Reference-Augmented Training for ASV Anti-Spoofing

DGX agent

arXiv:2606.10908v1 Announce Type: cross Abstract: We introduce a spoofing countermeasure architecture conditioned on speaker-reference recordings, but observe that it converges to a solution that effe

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

READER: Robust Evidence-based Authorship Decoding via Extracted Representations

DGX agent

arXiv:2606.10794v1 Announce Type: new Abstract: As agentic applications increasingly route user tasks through official and third-party LLM APIs, provenance becomes an operational question: which model

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

REAL: A Reasoning-Enhanced Graph Framework for Long-Term Memory Management of LLMs

DGX agent

arXiv:2606.10694v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly expected to interact with users over long time horizons. However, due to their finite context window, LLMs

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

RealMath-Eval: Why SOTA Judges Struggle with Real Human Reasoning

DGX agent

arXiv:2606.10254v1 Announce Type: new Abstract: While Large Language Models (LLMs) have achieved near-perfect performance in solving high-school mathematics, their ability to evaluate the diverse reas

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models

DGX agent

arXiv:2606.11164v1 Announce Type: new Abstract: Long chain-of-thought (CoT) trajectories in large language model (LLM) reasoning cause severe inference bottlenecks due to rapid key-value (KV) cache gr

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Recalling Too Well: Sycophancy Evaluation and Mitigation in Memory-Augmented Models

DGX agent

arXiv:2606.10949v1 Announce Type: new Abstract: Persistent memory systems promise to make LLMs more helpful by storing user beliefs over time. We show they also make models less correct by systematica

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Recoverable but Not Stationary:Local Linear Structures in Weights and Activations

DGX agent

arXiv:2606.10929v1 Announce Type: cross Abstract: Task vectors, LoRA, activation steering, and random search around pretrained weights all suggest that learned behaviour can be controlled by linear di

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

RedAct: Redacting Agent Capability Traces for Procedural Skill Protection

DGX agent

arXiv:2606.10813v1 Announce Type: cross Abstract: Users rely on execution traces to observe agent behavior, diagnose failures, and ensure accountability. These traces contain rich procedural detail, i

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

ReflectiChain: Epistemic Grounding in LLM-Driven World Models for Supply Chain Resilience

DGX agent

arXiv:2606.10359v1 Announce Type: new Abstract: AI agents in supply chains face a fundamental epistemic gap: large language models (LLMs) interpret policies but lack physical grounding, while reinforc

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Revisiting Metric Reliability for Fine-grained Evaluation of Machine Translation and Summarization in Indian Languages

DGX agent

arXiv:2510.07061v2 Announce Type: replace Abstract: While automatic metrics drive progress in Machine Translation (MT) and Text Summarization (TS), existing metrics have been developed and validated a

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning

DGX agent

arXiv:2510.14828v3 Announce Type: replace Abstract: Improving the reasoning capabilities of embodied agents is crucial for robots to complete complex human instructions in long-view manipulation tasks

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

RobotEQ: Transitioning from Passive Intelligence to Active Intelligence in Embodied AI

DGX agent

arXiv:2605.06234v2 Announce Type: replace Abstract: Embodied AI is a prominent research topic in both academia and industry. Current research centers on completing tasks based on explicit user instruc

model-releasesarxiv-cs-ro
10 Jun 2026
Model Releases

Rotate2Think: Geometric Priming via Orthogonal Rotation to Improve Language Model Reasoning

DGX agent

arXiv:2606.09873v1 Announce Type: cross Abstract: Reasoning models achieve strong performance on challenging tasks by generating explicit intermediate reasoning traces before producing a final answer.

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SAFE: An LLM-as-Verifier Framework for Evidence-Grounded Multi-Hop Reasoning

DGX agent

arXiv:2604.01993v2 Announce Type: replace-cross Abstract: Multi-hop QA benchmarks often reward Large Language Models (LLMs) for spurious correctness, where models reach correct answers through invalid

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Sample Where You Struggle: Sharpening Base Model Reasoning via Entropy-Guided Power Sampling

DGX agent

arXiv:2606.09926v1 Announce Type: cross Abstract: Sampling from the sequence-level power distribution p^alpha elicits RL-level reasoning from base language models without any parameter updates, but th

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation

DGX agent

arXiv:2606.10305v1 Announce Type: new Abstract: Fine-tuning vision-language-action (VLA) policies for long-horizon manipulation still relies heavily on behavior cloning, which requires costly high-qua

model-releasesarxiv-cs-ro
10 Jun 2026
Model Releases

SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning

DGX agent

arXiv:2606.10804v1 Announce Type: new Abstract: Controlled character animation requires transferring motion from a driving sequence to a reference character. Prior works heavily rely on intermediate r

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

SCOPE: Sequential Causal Optimization of Process Interventions

DGX agent

arXiv:2512.17629v4 Announce Type: replace-cross Abstract: Prescriptive Process Monitoring (PresPM) recommends interventions during running business processes to optimize key performance indicators (KP

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SHAPE: Coalition-Aware Expert Pruning for Sparse Mixture-of-Experts LLMs

DGX agent

arXiv:2606.09886v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (MoE) large language models achieve strong quality with low per-token compute, yet their deployment is often limited by the

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SHAPO: Sharpness-Aware Policy Optimization for Safe Exploration

DGX agent

arXiv:2606.10228v1 Announce Type: cross Abstract: Safe exploration is a prerequisite for deploying reinforcement learning (RL) agents in safety-critical domains. In this paper, we approach safe explor

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Sigma-Branch: Hierarchical Single-Path Network Reconstruction for Dynamic Inference with Reduced Active Parameters

DGX agent

arXiv:2606.09924v1 Announce Type: cross Abstract: Deploying deep neural networks on memory-constrained edge accelerators is bottlenecked by per-inference off-chip weight transfer rather than computati

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Sim2Schedule: A Simulator-Guided LLM Framework for Autonomous Open-Pit Mine Scheduling

DGX agent

arXiv:2606.10286v1 Announce Type: new Abstract: Open-pit mine scheduling is a critical process for maximizing economic return under complex geotechnical and operational constraints. While Mixed-Intege

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SkillResolve-Bench: Measuring and Resolving Same-Capability Ambiguity in Agent Skill Retrieval

DGX agent

arXiv:2606.10388v1 Announce Type: cross Abstract: Agent skill libraries are becoming routable software assets: a retrieved skill can contribute instructions, scripts, resource bindings, and execution

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Small Data, Big Noise: Adversarial Training for Robust Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2606.10610v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) has become essential for adapting foundation models to downstream NLP tasks. However, current PEFT methods often

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

SPACE: Source-free Proxy Anchor Concept Erasure for MLLMs

DGX agent

arXiv:2606.09868v1 Announce Type: cross Abstract: As Multimodal Large Language Models (MLLMs) face growing privacy risks and regulatory constraints, machine unlearning (MU) has emerged as a crucial so

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SPDM: Geometry-Modulated State Space Modeling with Manifold Constraints for Time Series Forecasting

DGX agent

arXiv:2606.09917v1 Announce Type: new Abstract: Multivariate time series forecasting requires capturing the continuously evolving correlation structure among interacting variables. Existing state-spac

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

SpineReport: Automated 3D Quantification and Reporting of Lumbar Spine Degeneration on MRI

DGX agent

arXiv:2606.10021v1 Announce Type: new Abstract: Lumbar spine conditions are a leading cause of disability worldwide, yet reliable quantification of degeneration from MRI remains challenging. In clinic

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

SSR-Merge: Subspace Signal Routing for Training-Free LoRA Merging in Diffusion Models

DGX agent

arXiv:2606.10617v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) merging can efficiently combine diverse generative capabilities from multiple trained LoRAs for a diffusion model. However, e

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

STAGE-Claw: Automated State-based Agent Benchmarking for Realistic Scenarios

DGX agent

arXiv:2606.10394v1 Announce Type: new Abstract: Large language models are increasingly used to power personal agents for everyday applications, but evaluating these agents remains a challenge. Existin

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Streaming Knowledge Compilation: Proactive Materiality-Scored Pinning for Time-Evolving LLM Wikis

DGX agent

arXiv:2606.09877v1 Announce Type: cross Abstract: LLM wiki systems compile knowledge into pre-filled KV caches for efficient inference, but assume a static corpus -- an assumption that fails whenever

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Structure from Reasoning, Numbers from Search: On-Premise Open LLMs as Structural Priors for Coupled MIMO Controller Tuning

DGX agent

arXiv:2606.11015v1 Announce Type: new Abstract: Tuning controllers for strongly coupled multi-input multi-output (MIMO) industrial processes is hard: decentralized classical auto-tuning ignores loop i

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

T1-Bench: Benchmarking Multi-Scenario Agents in Real-World Domains

DGX agent

arXiv:2606.11070v1 Announce Type: cross Abstract: Recent advances in reasoning and tool-calling capabilities of large language models (LLMs) have enabled increasingly capable agentic systems. However,

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Temporal Context Conditioning for Seasonality-Aware Precipitation Nowcasting of High-Intensity Rainfall

DGX agent

arXiv:2606.09959v1 Announce Type: cross Abstract: Precipitation nowcasting is increasingly being approached with deep learning models that learn directly from recent radar observations. Although such

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Temporal Sheaf Neural Networks with Dynamic Orthogonal Transport

DGX agent

arXiv:2606.10071v1 Announce Type: cross Abstract: We introduce Temporal Sheaf Neural Networks (TSNN), a temporal link prediction framework that equips each node with a time-varying orthogonal frame an

model-releasesarxiv-cs-ai
10 Jun 2026
← Previous
1…138139140141142…361
Next →