AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,805 results
Model Releases

Recognizing Co-Speech Gestures in-the-Wild

DGX agent

arXiv:2605.31589v1 Announce Type: new Abstract: While humans naturally gesture during speech, only a sparse subset of these movements are visually depictive and semantically linked to specific spoken

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

ReTabAD: A Benchmark for Restoring Semantic Context in Tabular Anomaly Detection

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2510.02060v2 Announce Type: replace Abstract: In tabular anomaly detection (AD), textual semantics often carry critical signals, as the definition of an anomaly is closely tied to domain-specifi

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

SAEmnesia: Erasing Concepts in Diffusion Models with Supervised Sparse Autoencoders

DGX agent

arXiv:2509.21379v3 Announce Type: replace-cross Abstract: Concept unlearning in diffusion models is hampered by feature splitting, where concepts are distributed across many latent features, making th

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Safe Equilibrium Policy Optimization for Strategic Agent Policies

DGX agent

arXiv:2605.30854v1 Announce Type: cross Abstract: Language models fine-tuned with reinforcement learning typically optimize for task reward, ignoring multi-agent strategic structure. Because these age

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Same Patient, Different Words, Different Diagnosis? Evaluating Semantic Stability in Clinical LLMs

DGX agent

arXiv:2605.30646v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in clinical applications. However, their behavior remains highly sensitive to subtle linguistic var

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

SAW-Bench: Learning Situated Awareness in the Real World

DGX agent

arXiv:2602.16682v2 Announce Type: replace Abstract: A core aspect of human perception is situated awareness, the ability to relate ourselves to the surrounding physical environment and reason over pos

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Scaling Conversational Hungarian ASR: The BEA-Dialogue+ Corpus

DGX agent

arXiv:2605.31469v1 Announce Type: cross Abstract: Conversational automatic speech recognition in Hungarian is constrained by the limited amount of publicly available dialogue-style training data. The

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Scaling Multi-Hop Training Data via Graph-Constrained Path Selection

DGX agent

arXiv:2605.31238v1 Announce Type: new Abstract: Endowing large language models with compositional reasoning over specialized documents requires multi-hop training data at scale, where such data rarely

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Self-Tuning Regularization for Image Scanning Microscopy

DGX agent

arXiv:2605.31426v1 Announce Type: cross Abstract: Image Scanning Microscopy (ISM) is a fluorescence imaging technique that combines detector-array acquisition and computational reconstruction to achie

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Send a SCOUT First: Pre-hoc Reasoning for Adaptive Detector Allocation in Prompt-Injection Defense

DGX agent

arXiv:2605.30837v1 Announce Type: cross Abstract: Prompt-injection detectors are heterogeneous: each is strong on a different slice of attacks, and none is always reliable. Yet existing systems still

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Sequential Least-Squares Estimators with Fast Randomized Sketching for Linear Statistical Models

DGX agent

arXiv:2509.06856v2 Announce Type: replace-cross Abstract: We propose a novel randomized framework for the estimation problem of large-scale linear statistical models, namely Sequential Least-Squares E

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Sequential Subspace Noise Injection Prevents Accuracy Collapse in Certified Unlearning

DGX agent

arXiv:2601.05134v2 Announce Type: replace Abstract: Certified unlearning based on differential privacy offers strong guarantees but remains largely impractical: the noisy fine-tuning approaches propos

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

SERA: Soft-Verified Efficient Repository Agents

DGX agent

arXiv:2601.20789v3 Announce Type: replace Abstract: Open-weight coding agents should hold a fundamental advantage over closed-source systems because they can specialize to private codebases, encoding

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

SimulCost: A Cost-Aware Benchmark and Toolkit for Automating Physics Simulations with LLMs

DGX agent

arXiv:2603.20253v2 Announce Type: replace-cross Abstract: Evaluating LLM agents for scientific tasks has focused on token costs while ignoring tool-use costs like simulation time and experimental reso

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Skill Availability and Presentation Granularity in Large-Language-Model Agents: A Controlled SkillsBench Study

DGX agent

arXiv:2605.31408v1 Announce Type: cross Abstract: Skill documents provide procedural knowledge to large-language-model agents at inference time. This article studies whether the presentation granulari

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Smaller and Faster 3DGS via Post-Training Dictionary Learning

DGX agent

arXiv:2605.30396v1 Announce Type: cross Abstract: 3D Gaussian Splatting (3DGS) is a promising neural scene representation for real-time rendering, but trained models often suffer from large memory foo

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

So much great work lately from Nvidia, the 'King of American Open-source AI'! - Crossed 1,000 total public repositories on @huggingface (820…

DGX agent

So much great work lately from Nvidia, the 'King of American Open-source AI'! - Crossed 1,000 total public repositories on @huggingface (820 models, 249 datasets & 57 spaces) & almost 60,000 followers

model-releasesclem-delangue--x
1 Jun 2026
Model Releases

Social welfare optimisation under institutional reward and punishment

DGX agent

arXiv:2605.31330v1 Announce Type: cross Abstract: Institutional incentives are widely used to promote cooperation among autonomous, self-regarding agents, from human societies to multi-agent and AI sy

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

SOCO: Benchmarking Semantic Object Correspondence in Vision Foundation Models

DGX agent

arXiv:2605.31597v1 Announce Type: new Abstract: Measuring structured object understanding in vision foundation models remains challenging due to inconsistent evaluation protocols and limited part-leve

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Softsign: Smooth Sign in Your Optimizer For Better Parameter Heterogeneity Handling

DGX agent

arXiv:2605.31371v1 Announce Type: new Abstract: Sign-based and LMO-inspired optimizers have recently attracted substantial attention in deep learning due to their strong performance and low memory foo

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

SpatialAct: Probing Spatial Reasoning-to-Action Capabilities of VLM Agents in 3D Scenes

DGX agent

arXiv:2605.31148v1 Announce Type: cross Abstract: Humans can effortlessly perceive spatial layouts, form cognitive representations, reason about spatial relations, and translate such reasoning into ac

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

SPM-Bench: Benchmarking Large Language Models for Scanning Probe Microscopy

DGX agent

arXiv:2602.22971v2 Announce Type: replace Abstract: As LLMs achieved breakthroughs in general reasoning, their proficiency in specialized scientific domains reveals pronounced gaps in existing benchma

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Steering LLMs? Actually, Sparse Autoencoders can outperform simple baselines

DGX agent

arXiv:2605.31183v1 Announce Type: cross Abstract: Sparse Autoencoders (SAEs) have been seen as a promising avenue for exploring the internals of Large Language Models (LLMs) and for steering model out

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Structured interactions improve distributed coordination beyond model scaling in a real-world multi-robot system

DGX agent

arXiv:2605.30383v1 Announce Type: cross Abstract: Scaling individual robot capabilities is common but costly. Here we investigate a system-level design question in real-world multi-robot coordination:

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

SVI-Bench: A Dynamic Microworld for Strategic Video Intelligence

DGX agent

arXiv:2605.31529v1 Announce Type: new Abstract: True video intelligence demands more than recognizing what is visible: it requires reasoning about why events unfold, predicting what would change under

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Symbolic Intermediaries as a Linguistic-Numerical Interface for LLM-Driven Geometric Reasoning

DGX agent

arXiv:2505.17607v3 Announce Type: replace Abstract: Large Language Models (LLMs) display reasoning capabilities over linguistic and symbolic objects but have limited capabilities to directly interpret

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

TabCausal: Pretraining Across Causal Environments for Tabular Causal Discovery

DGX agent

arXiv:2605.31156v1 Announce Type: new Abstract: Causal discovery aims to recover directed causal relations from observational and interventional data, providing a basis for mechanistic understanding a

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

TAGA: A Tangent-Based Reactive Approach for Socially Compliant Robot Navigation Around Human Groups

DGX agent

arXiv:2503.21168v3 Announce Type: replace Abstract: Robots navigating human-populated environments must avoid collisions while respecting the social structure of crowds, particularly the implicit boun

model-releasesarxiv-cs-ro
1 Jun 2026
Model Releases

Target-Agnostic Calibration under Distribution Shift with Frequency-Aware Gradient Rectification

DGX agent

arXiv:2508.19830v2 Announce Type: replace-cross Abstract: Real-world model deployments inevitably encounter distribution shifts, rendering the confidence estimates of deep neural networks highly unrel

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Targeted Speaker Poisoning Framework in Zero-Shot Text-to-Speech

DGX agent

arXiv:2603.07551v2 Announce Type: replace-cross Abstract: Zero-shot Text-to-Speech (TTS) voice cloning poses severe privacy risks, demanding the removal of specific speaker identities from trained TTS

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

TaxoBell: Gaussian Box Embeddings for Self-Supervised Taxonomy Expansion

DGX agent

arXiv:2601.09633v2 Announce Type: replace Abstract: Taxonomies form the backbone of structured knowledge representation across diverse domains, enabling applications such as e-commerce and semantic se

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

TeachObs: A Human-Validated Benchmark for Multimodal Teaching Observation and Model Evaluation

DGX agent

arXiv:2605.30673v1 Announce Type: new Abstract: Classroom videos contain observable teaching practices, but their pedagogical and visual signals are rarely organized in forms suitable for model evalua

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

The fully-managed Remote MCP Server for AlloyDB is now Generally Available

DGX agent

AI agents possess incredible reasoning capabilities and can perform increasingly complex actions. But the reliability of agentic outcomes depends entirely on the quality of the context they can access

model-releasesgoogle-cloud-ai
1 Jun 2026
Model Releases

The Geometry of Activity Cliffs: Representation Dependence and Multi-Scale Characterization of Activity Landscapes

DGX agent

arXiv:2605.30831v1 Announce Type: cross Abstract: Activity cliffs, structurally similar compounds with large potency differences, are widely treated as intrinsic features of chemical datasets. We argu

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

The Illusion of Generalization in Tabular Language Models

DGX agent

arXiv:2602.04031v2 Announce Type: replace Abstract: Tabular Language Models (TLMs) have been claimed to achieve strong generalization for tabular prediction. We conduct a systematic re-evaluation of T

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

The Regularizing Power of Language-Training Deepfake Detectors

DGX agent

arXiv:2605.31192v1 Announce Type: new Abstract: Recently, thanks to the advent of Multimodal-LLMs, deepfake detectors are striving not only to be generalizable but also interpretable. We propose that

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

The Surface You Test Is Not the Surface That Breaks

DGX agent

arXiv:2605.30454v1 Announce Type: cross Abstract: Tool-augmented LLM agents are vulnerable to prompt injection: a third party who controls part of the agent's context can plant instructions that the a

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Thinking in Structures: Evaluating Spatial Intelligence in Constraint-Governed Spaces

DGX agent

arXiv:2602.07864v2 Announce Type: replace Abstract: Spatial intelligence is crucial for vision--language models (VLMs), yet many scene-centric benchmarks evaluate unconstrained environments where a si

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

TRACE: Discovering Task-Specific Parameter via Adaptation-Aware Probing for Continual Fine-Tuning

DGX agent

arXiv:2605.31025v1 Announce Type: new Abstract: In real-world deployment, LLMs are often adapted continually across tasks to keep LLMs up-to-date in production, where new fine-tuning should preserve p

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

TraceGraph: Shared Decision Landscapes for Diagnosing and Improving Agent Trajectories

DGX agent

arXiv:2605.31308v1 Announce Type: new Abstract: Agent benchmarks increasingly record rich interaction trajectories, yet evaluation often reduces each rollout to a pass rate or reward score. We introdu

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Translation Analytics for Freelancers II: Benchmarking Local LLMs for Confidential Translation Workflows

DGX agent

arXiv:2605.31452v1 Announce Type: new Abstract: Building on our previous work, this paper develops practical, low-barrier methods for freelance translators and smaller language service providers to ev

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Triaging Threats to Specialized Guardrails

DGX agent

arXiv:2605.30693v1 Announce Type: cross Abstract: Building robust safety guardrails is essential for deploying Large Language Models across diverse real-world applications. However, this goal remains

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

TSM-Bench: Detecting LLM-Generated Text in Real-World Wikipedia Editing Practices

DGX agent

arXiv:2605.31113v1 Announce Type: new Abstract: Automatically detecting machine-generated text (MGT) is critical to maintaining the knowledge integrity of user-generated content (UGC) platforms such a

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

UniDial-EvalKit: A Unified Toolkit for Evaluating Multi-Faceted Conversational Abilities

DGX agent

arXiv:2603.23160v2 Announce Type: replace Abstract: Benchmarking large language models (LLMs) and agents in multi-turn interactive scenarios is essential for understanding their practical capabilities

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Unmute the Patch Tokens: Rethinking Probing in Multi-Label Audio Classification

DGX agent

arXiv:2509.24901v4 Announce Type: replace-cross Abstract: Although probing frozen models has become a standard evaluation paradigm, self-supervised learning in audio defaults to fine-tuning when pursu

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Unraveling LoRA Interference: Orthogonal Subspaces for Robust Model Merging

DGX agent

arXiv:2505.22934v2 Announce Type: replace-cross Abstract: Fine-tuning large language models (LMs) for individual tasks yields strong performance but is expensive for deployment and storage. Recent wor

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Variational Routing: A Scalable Bayesian Framework for Calibrated Mixture-of-Experts Transformers

DGX agent

arXiv:2603.09453v3 Announce Type: replace-cross Abstract: Foundation models are increasingly being deployed in contexts where understanding the uncertainty of their outputs is critical to ensuring res

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Very good advice on self-improving agents. (bookmark it) This is something I am seeing in my own experiments with coding agents and harnesse…

DGX agent

Very good advice on self-improving agents. (bookmark it) This is something I am seeing in my own experiments with coding agents and harnesses for long-horizon tasks. What I have found is that stronger

model-releasesdair-ai--x
1 Jun 2026
← Previous
1…242243244245246…476
Next →