AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Re-examining Low Rank adaptation for private LLM fine-tuning

DGX agent

arXiv:2510.01137v3 Announce Type: replace Abstract: Privacy is a central concern when fine-tuning large language models (LLMs) on sensitive data, and differentially private stochastic gradient descent

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Recognizing Co-Speech Gestures in-the-Wild

DGX agent

arXiv:2605.31589v1 Announce Type: new Abstract: While humans naturally gesture during speech, only a sparse subset of these movements are visually depictive and semantically linked to specific spoken

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

ReTabAD: A Benchmark for Restoring Semantic Context in Tabular Anomaly Detection

DGX agent

arXiv:2510.02060v2 Announce Type: replace Abstract: In tabular anomaly detection (AD), textual semantics often carry critical signals, as the definition of an anomaly is closely tied to domain-specifi

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

SAEmnesia: Erasing Concepts in Diffusion Models with Supervised Sparse Autoencoders

DGX agent

arXiv:2509.21379v3 Announce Type: replace-cross Abstract: Concept unlearning in diffusion models is hampered by feature splitting, where concepts are distributed across many latent features, making th

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Safe Equilibrium Policy Optimization for Strategic Agent Policies

DGX agent

arXiv:2605.30854v1 Announce Type: cross Abstract: Language models fine-tuned with reinforcement learning typically optimize for task reward, ignoring multi-agent strategic structure. Because these age

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Same Patient, Different Words, Different Diagnosis? Evaluating Semantic Stability in Clinical LLMs

DGX agent

arXiv:2605.30646v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in clinical applications. However, their behavior remains highly sensitive to subtle linguistic var

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

SAW-Bench: Learning Situated Awareness in the Real World

DGX agent

arXiv:2602.16682v2 Announce Type: replace Abstract: A core aspect of human perception is situated awareness, the ability to relate ourselves to the surrounding physical environment and reason over pos

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Scaling Conversational Hungarian ASR: The BEA-Dialogue+ Corpus

DGX agent

arXiv:2605.31469v1 Announce Type: cross Abstract: Conversational automatic speech recognition in Hungarian is constrained by the limited amount of publicly available dialogue-style training data. The

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Scaling Multi-Hop Training Data via Graph-Constrained Path Selection

DGX agent

arXiv:2605.31238v1 Announce Type: new Abstract: Endowing large language models with compositional reasoning over specialized documents requires multi-hop training data at scale, where such data rarely

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Self-Tuning Regularization for Image Scanning Microscopy

DGX agent

arXiv:2605.31426v1 Announce Type: cross Abstract: Image Scanning Microscopy (ISM) is a fluorescence imaging technique that combines detector-array acquisition and computational reconstruction to achie

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Send a SCOUT First: Pre-hoc Reasoning for Adaptive Detector Allocation in Prompt-Injection Defense

DGX agent

arXiv:2605.30837v1 Announce Type: cross Abstract: Prompt-injection detectors are heterogeneous: each is strong on a different slice of attacks, and none is always reliable. Yet existing systems still

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Sequential Least-Squares Estimators with Fast Randomized Sketching for Linear Statistical Models

DGX agent

arXiv:2509.06856v2 Announce Type: replace-cross Abstract: We propose a novel randomized framework for the estimation problem of large-scale linear statistical models, namely Sequential Least-Squares E

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Sequential Subspace Noise Injection Prevents Accuracy Collapse in Certified Unlearning

DGX agent

arXiv:2601.05134v2 Announce Type: replace Abstract: Certified unlearning based on differential privacy offers strong guarantees but remains largely impractical: the noisy fine-tuning approaches propos

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

SERA: Soft-Verified Efficient Repository Agents

DGX agent

arXiv:2601.20789v3 Announce Type: replace Abstract: Open-weight coding agents should hold a fundamental advantage over closed-source systems because they can specialize to private codebases, encoding

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

SimulCost: A Cost-Aware Benchmark and Toolkit for Automating Physics Simulations with LLMs

DGX agent

arXiv:2603.20253v2 Announce Type: replace-cross Abstract: Evaluating LLM agents for scientific tasks has focused on token costs while ignoring tool-use costs like simulation time and experimental reso

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Skill Availability and Presentation Granularity in Large-Language-Model Agents: A Controlled SkillsBench Study

DGX agent

arXiv:2605.31408v1 Announce Type: cross Abstract: Skill documents provide procedural knowledge to large-language-model agents at inference time. This article studies whether the presentation granulari

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Smaller and Faster 3DGS via Post-Training Dictionary Learning

DGX agent

arXiv:2605.30396v1 Announce Type: cross Abstract: 3D Gaussian Splatting (3DGS) is a promising neural scene representation for real-time rendering, but trained models often suffer from large memory foo

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Social welfare optimisation under institutional reward and punishment

DGX agent

arXiv:2605.31330v1 Announce Type: cross Abstract: Institutional incentives are widely used to promote cooperation among autonomous, self-regarding agents, from human societies to multi-agent and AI sy

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

SOCO: Benchmarking Semantic Object Correspondence in Vision Foundation Models

DGX agent

arXiv:2605.31597v1 Announce Type: new Abstract: Measuring structured object understanding in vision foundation models remains challenging due to inconsistent evaluation protocols and limited part-leve

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Softsign: Smooth Sign in Your Optimizer For Better Parameter Heterogeneity Handling

DGX agent

arXiv:2605.31371v1 Announce Type: new Abstract: Sign-based and LMO-inspired optimizers have recently attracted substantial attention in deep learning due to their strong performance and low memory foo

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

SpatialAct: Probing Spatial Reasoning-to-Action Capabilities of VLM Agents in 3D Scenes

DGX agent

arXiv:2605.31148v1 Announce Type: cross Abstract: Humans can effortlessly perceive spatial layouts, form cognitive representations, reason about spatial relations, and translate such reasoning into ac

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

SPM-Bench: Benchmarking Large Language Models for Scanning Probe Microscopy

DGX agent

arXiv:2602.22971v2 Announce Type: replace Abstract: As LLMs achieved breakthroughs in general reasoning, their proficiency in specialized scientific domains reveals pronounced gaps in existing benchma

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Steering LLMs? Actually, Sparse Autoencoders can outperform simple baselines

DGX agent

arXiv:2605.31183v1 Announce Type: cross Abstract: Sparse Autoencoders (SAEs) have been seen as a promising avenue for exploring the internals of Large Language Models (LLMs) and for steering model out

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Structured interactions improve distributed coordination beyond model scaling in a real-world multi-robot system

DGX agent

arXiv:2605.30383v1 Announce Type: cross Abstract: Scaling individual robot capabilities is common but costly. Here we investigate a system-level design question in real-world multi-robot coordination:

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

SVI-Bench: A Dynamic Microworld for Strategic Video Intelligence

DGX agent

arXiv:2605.31529v1 Announce Type: new Abstract: True video intelligence demands more than recognizing what is visible: it requires reasoning about why events unfold, predicting what would change under

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Symbolic Intermediaries as a Linguistic-Numerical Interface for LLM-Driven Geometric Reasoning

DGX agent

arXiv:2505.17607v3 Announce Type: replace Abstract: Large Language Models (LLMs) display reasoning capabilities over linguistic and symbolic objects but have limited capabilities to directly interpret

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

TabCausal: Pretraining Across Causal Environments for Tabular Causal Discovery

DGX agent

arXiv:2605.31156v1 Announce Type: new Abstract: Causal discovery aims to recover directed causal relations from observational and interventional data, providing a basis for mechanistic understanding a

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

TAGA: A Tangent-Based Reactive Approach for Socially Compliant Robot Navigation Around Human Groups

DGX agent

arXiv:2503.21168v3 Announce Type: replace Abstract: Robots navigating human-populated environments must avoid collisions while respecting the social structure of crowds, particularly the implicit boun

model-releasesarxiv-cs-ro
1 Jun 2026
Model Releases

Target-Agnostic Calibration under Distribution Shift with Frequency-Aware Gradient Rectification

DGX agent

arXiv:2508.19830v2 Announce Type: replace-cross Abstract: Real-world model deployments inevitably encounter distribution shifts, rendering the confidence estimates of deep neural networks highly unrel

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Targeted Speaker Poisoning Framework in Zero-Shot Text-to-Speech

DGX agent

arXiv:2603.07551v2 Announce Type: replace-cross Abstract: Zero-shot Text-to-Speech (TTS) voice cloning poses severe privacy risks, demanding the removal of specific speaker identities from trained TTS

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

TaxoBell: Gaussian Box Embeddings for Self-Supervised Taxonomy Expansion

DGX agent

arXiv:2601.09633v2 Announce Type: replace Abstract: Taxonomies form the backbone of structured knowledge representation across diverse domains, enabling applications such as e-commerce and semantic se

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

TeachObs: A Human-Validated Benchmark for Multimodal Teaching Observation and Model Evaluation

DGX agent

arXiv:2605.30673v1 Announce Type: new Abstract: Classroom videos contain observable teaching practices, but their pedagogical and visual signals are rarely organized in forms suitable for model evalua

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

The Geometry of Activity Cliffs: Representation Dependence and Multi-Scale Characterization of Activity Landscapes

DGX agent

arXiv:2605.30831v1 Announce Type: cross Abstract: Activity cliffs, structurally similar compounds with large potency differences, are widely treated as intrinsic features of chemical datasets. We argu

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

The Illusion of Generalization in Tabular Language Models

DGX agent

arXiv:2602.04031v2 Announce Type: replace Abstract: Tabular Language Models (TLMs) have been claimed to achieve strong generalization for tabular prediction. We conduct a systematic re-evaluation of T

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

The Regularizing Power of Language-Training Deepfake Detectors

DGX agent

arXiv:2605.31192v1 Announce Type: new Abstract: Recently, thanks to the advent of Multimodal-LLMs, deepfake detectors are striving not only to be generalizable but also interpretable. We propose that

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

The Surface You Test Is Not the Surface That Breaks

DGX agent

arXiv:2605.30454v1 Announce Type: cross Abstract: Tool-augmented LLM agents are vulnerable to prompt injection: a third party who controls part of the agent's context can plant instructions that the a

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Thinking in Structures: Evaluating Spatial Intelligence in Constraint-Governed Spaces

DGX agent

arXiv:2602.07864v2 Announce Type: replace Abstract: Spatial intelligence is crucial for vision--language models (VLMs), yet many scene-centric benchmarks evaluate unconstrained environments where a si

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

TRACE: Discovering Task-Specific Parameter via Adaptation-Aware Probing for Continual Fine-Tuning

DGX agent

arXiv:2605.31025v1 Announce Type: new Abstract: In real-world deployment, LLMs are often adapted continually across tasks to keep LLMs up-to-date in production, where new fine-tuning should preserve p

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

TraceGraph: Shared Decision Landscapes for Diagnosing and Improving Agent Trajectories

DGX agent

arXiv:2605.31308v1 Announce Type: new Abstract: Agent benchmarks increasingly record rich interaction trajectories, yet evaluation often reduces each rollout to a pass rate or reward score. We introdu

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Translation Analytics for Freelancers II: Benchmarking Local LLMs for Confidential Translation Workflows

DGX agent

arXiv:2605.31452v1 Announce Type: new Abstract: Building on our previous work, this paper develops practical, low-barrier methods for freelance translators and smaller language service providers to ev

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Triaging Threats to Specialized Guardrails

DGX agent

arXiv:2605.30693v1 Announce Type: cross Abstract: Building robust safety guardrails is essential for deploying Large Language Models across diverse real-world applications. However, this goal remains

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

TSM-Bench: Detecting LLM-Generated Text in Real-World Wikipedia Editing Practices

DGX agent

arXiv:2605.31113v1 Announce Type: new Abstract: Automatically detecting machine-generated text (MGT) is critical to maintaining the knowledge integrity of user-generated content (UGC) platforms such a

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

UniDial-EvalKit: A Unified Toolkit for Evaluating Multi-Faceted Conversational Abilities

DGX agent

arXiv:2603.23160v2 Announce Type: replace Abstract: Benchmarking large language models (LLMs) and agents in multi-turn interactive scenarios is essential for understanding their practical capabilities

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Unmute the Patch Tokens: Rethinking Probing in Multi-Label Audio Classification

DGX agent

arXiv:2509.24901v4 Announce Type: replace-cross Abstract: Although probing frozen models has become a standard evaluation paradigm, self-supervised learning in audio defaults to fine-tuning when pursu

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Unraveling LoRA Interference: Orthogonal Subspaces for Robust Model Merging

DGX agent

arXiv:2505.22934v2 Announce Type: replace-cross Abstract: Fine-tuning large language models (LMs) for individual tasks yields strong performance but is expensive for deployment and storage. Recent wor

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Variational Routing: A Scalable Bayesian Framework for Calibrated Mixture-of-Experts Transformers

DGX agent

arXiv:2603.09453v3 Announce Type: replace-cross Abstract: Foundation models are increasingly being deployed in contexts where understanding the uncertainty of their outputs is critical to ensuring res

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Weight Decay Improves Language Model Plasticity

DGX agent

arXiv:2602.11137v2 Announce Type: replace-cross Abstract: Large language models are typically trained in two broad phases: pretraining to produce a base model, followed by further training to improve

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Welfare, Improvability, and Variance: A Principal-Agent Approach to Optimal Benchmark Item Aggregation

DGX agent

arXiv:2605.30916v1 Announce Type: new Abstract: AI benchmarks have well-documented limitations, with prior work examining contamination, saturation, and construct underspecification. Aggregation has r

model-releasesarxiv-cs-lg
1 Jun 2026
← Previous
1…177178179180181…361
Next →