AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,603 results
9 Jun 2026

Distilling Safe LLM Systems via Soft Prompts for On Device Settings

Model ReleasesDGX agent

arXiv:2606.09388v1 Announce Type: new Abstract: Deploying safe large language models (LLMs) on resource-constrained edge devices presents a critical challenge: while dual-model systems combining LLMs

Distortion-Aware PETR for BEV Object Detection with Mixed Pinhole-Fisheye Cameras

Model ReleasesDGX agent

arXiv:2606.08680v1 Announce Type: new Abstract: Fisheye cameras are widely deployed in autonomous driving perception suites for their low cost and full-coverage field of view (FOV), yet their potentia

DIYHealth Suite: Dataset, Model, and Benchmark for Health Management at Home

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.07542v1 Announce Type: cross Abstract: Generative AI is reshaping healthcare, yet most existing advances rely on hospital-grade devices, which limits their accessibility and potential for h

Domain-Adapted Small Language Models with Hybrid Post-Processing: Achieving Cost-Efficient, Low-Latency Multi-Label Structured Prediction via LoRA Fine-Tuning on Scarce Data

Model ReleasesDGX agent

arXiv:2606.05781v2 Announce Type: replace Abstract: Deploying frontier large language models (LLMs) for domain-specific structured evaluation tasks incurs prohibitive latency, cost, and data-privacy o

Don’t always agree with @mustafasuleyman but in this case he is right to call out @AnthropicAI.

Model ReleasesDGX agent

Don’t always agree with @mustafasuleyman but in this case he is right to call out @AnthropicAI. Microsoft AI head calls out Anthropic for acting like Claude is conscious https://www.theverge.com/tech/

DriveReward: A Comprehensive Dataset and Generative Vision-Language Reward Model for Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.08525v1 Announce Type: new Abstract: Reward models play a pivotal role in reinforcement learning (RL) and multi-modal trajectory selection for autonomous driving. However, acquiring such re

Driving Video Retrieval for Complex Queries with Structured Grounding

Model ReleasesDGX agent

arXiv:2606.09109v1 Announce Type: new Abstract: Video retrieval at scale is central to data curation and safety validation in autonomous driving, where users want to find not only scenes but also dyna

Dual Quaternion-Based Unscented Kalman Filter with Visual Inertial Odometry for Navigation in GPS-Denied Environments

Model ReleasesDGX agent

arXiv:2606.09292v1 Announce Type: new Abstract: Reliable navigation in GPS-denied environments remains a fundamental challenge in robotics, aerospace, and autonomous vehicle applications. This paper p

Echo-DM: Ultrasound Marker Removal via Conditional Latent Diffusion and Region-Aware Fusion

Model ReleasesDGX agent

arXiv:2606.09378v1 Announce Type: new Abstract: Clinical ultrasound images often contain artificial markers, such as measurement calipers and text, to assist diagnostic interpretation and comparison.

Efficient Minimal Solvers for Relative Pose Estimation in Autonomous Driving Applications

Model ReleasesDGX agent

arXiv:2606.09569v1 Announce Type: cross Abstract: With the advancement of visual sensing systems, computer vision is playing an increasingly important role in autonomous driving and robot navigation.

Efficient Minimal Solvers for Visual-Inertial Relative Pose Estimation in Multi-Camera Systems

Model ReleasesDGX agent

arXiv:2606.09477v1 Announce Type: new Abstract: Estimating the relative poses of multi-camera systems is a fundamental problem in computer vision, with critical applications in autonomous vehicles, mo

EgoTactile: Learning Grasp Pressure for Everyday Objects from Egocentric Video

Model ReleasesDGX agent

arXiv:2606.09243v1 Announce Type: cross Abstract: Estimating full-hand grasp pressure from egocentric video is critical for immersive VR and robotic manipulation, yet dense tactile sensing often relie

Emergence World: A Platform for Evaluating Long-Horizon Multi-Agent Autonomy

Model ReleasesDGX agent

arXiv:2606.08367v1 Announce Type: cross Abstract: Most evaluations of LLM agents look like exams: a discrete task, a clean environment, a score in minutes or hours. We argue that this approach is mism

End-to-End Context Compression at Scale

Model ReleasesDGX agent

arXiv:2606.09659v1 Announce Type: cross Abstract: Long-context language model inference is bottlenecked by memory, as the KV cache grows with context length. Recent techniques to compress the KV cache

End-to-End Training for Discrete Token LLM based TTS System

Model ReleasesDGX agent

arXiv:2606.09234v1 Announce Type: cross Abstract: Recent state-of-the-art (SOTA) text-to-speech (TTS) systems typically adopt a cascaded pipeline consisting of a speech tokenizer, an autoregressive la

Enhanced Detection of Tiny Objects in Aerial Images

Model ReleasesDGX agent

arXiv:2509.17078v3 Announce Type: replace Abstract: While one-stage detectors like YOLOv8 offer fast training speed, they often under-perform on detecting small objects as a trade-off. This becomes ev

Enhancing Spatial Reasoning in Large Language Models for Metal-Organic Frameworks Structure Prediction

Model ReleasesDGX agent

arXiv:2601.09285v2 Announce Type: replace Abstract: Metal-organic frameworks (MOFs) are porous crystalline materials with broad applications such as carbon capture and drug delivery, yet accurately pr

ERBench: A Benchmark and Testsuite for Equation Discovery Algorithms

Model ReleasesDGX agent

arXiv:2606.09276v1 Announce Type: new Abstract: Equation discovery aims to automate the discovery of scientific models in the form of mathematical equations from data. Technically, equation discovery

Evaluate Clinical ASR Models Faster with Agent Skills and NVIDIA Nemotron Speech

Model ReleasesDGX agent

This article presents a clinical automatic speech recognition workflow for generating pronunciation-aware synthetic audio, reviewing clinical terms, and evaluating recognition quality using NVIDIA age

Evaluating Advanced Prompting on Gemini Flash for Multi-Hop Biomedical QA

Model ReleasesDGX agent

arXiv:2606.07548v1 Announce Type: cross Abstract: The MedHopQA challenge presents a critical test for Large Language Models (LLMs): complex, multi-hop reasoning in the high-stakes biomedical domain. T

Evaluating Hallucinations in Domain-Adapted Large Language Models

Model ReleasesDGX agent

arXiv:2606.07521v1 Announce Type: cross Abstract: This study investigates the phenomenon of hallucinations in domain-adapted Large Language Models (LLMs), focusing on the fine-tuning of the Llama-2 mo

Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting

Model ReleasesDGX agent

arXiv:2606.09809v1 Announce Type: new Abstract: AI evaluation results are produced at scale but reported inconsistently across leaderboards, model cards, benchmark papers, and company blogs. The cost

Executable World Models for ARC-AGI-3 in the Era of Coding Agents

Model ReleasesDGX agent

arXiv:2605.05138v2 Announce Type: replace Abstract: We evaluate an initial coding-agent system for ARC-AGI-3 in which the agent maintains an executable Python world model, verifies it against previous

Experience Makes Skillful: Enabling Generalizable Medical Agent Reasoning via Self-Evolving Skill Memory

Model ReleasesDGX agent

arXiv:2606.09365v1 Announce Type: new Abstract: Medical agent systems are increasingly expected to support interactive clinical decision making rather than only static question answering. In such sett

Explaining Black-Box Language Models: Learning to Optimize Linguistically-Structured Word Subsets

Model ReleasesDGX agent

arXiv:2606.08497v1 Announce Type: new Abstract: As deep language models (DLMs) are increasingly deployed in high-stakes domains such as healthcare, understanding their decision rationale becomes param

Exploring the Effect of Basis Rotation on NQS Performance

Model ReleasesDGX agent

arXiv:2512.17893v2 Announce Type: replace-cross Abstract: Neural Quantum States (NQS) are powerful variational representations of quantum many-body wavefunctions, yet their performance depends sensiti

Fable 5 is now available in Claude Code and Cowork Fable is the best model I have used for coding, by a wide margin. It is a big step up, en…

Model ReleasesDGX agent

Fable 5 is now available in Claude Code and Cowork Fable is the best model I have used for coding, by a wide margin. It is a big step up, enabling less prompts and steers, more efficient token use, be

🚨 Fable 5 is something to pay attention to. This is another 'I had early access to the new Mythos-class Anthropic model and I want to tell …

Model ReleasesDGX agent

🚨 Fable 5 is something to pay attention to. This is another 'I had early access to the new Mythos-class Anthropic model and I want to tell you what I thought of it' post. I know, I know, it's annoying

Fable 5 is the biggest step up I’ve felt in our models since Opus 4.5 back in November. After 4.5 came out I uninstalled my IDE when I reali…

Model ReleasesDGX agent

Fable 5 is the biggest step up I’ve felt in our models since Opus 4.5 back in November. After 4.5 came out I uninstalled my IDE when I realized that I’d been doing 100% of my coding in a terminal for

Fable is a step-change in models, and I hope it changes how you work with Claude. More to come in a series of posts on how it’s reshaped our…

Model ReleasesDGX agent

Fable is a step-change in models, and I hope it changes how you work with Claude. More to come in a series of posts on how it’s reshaped our work, but the TLDR: it’s time to be more ambitious. Claude

Families of Control-Cost-Parametrized Inverse-Optimal Universal Stabilizers

Model ReleasesDGX agent

arXiv:2606.09047v1 Announce Type: cross Abstract: A classical universal stabilization formula offers the practitioner no design freedom: it is a single, parameter-free object. We introduce a cost-para

Filigran launches XTM One to automate threat exposure management with AI agents

Model ReleasesDGX agent

French cybersecurity company Filigran SAS today launched XTM One, an artificial intelligence orchestration layer that automates continuous threat exposure management workflows across its platform. XTM

FineGen: A VLM-based Multi-Agent Framework for Fine-Grained Image-Text Dataset Construction

Model ReleasesDGX agent

arXiv:2606.07645v1 Announce Type: cross Abstract: The scarcity of hard negative samples in current vision-language datasets significantly hinders fine-grained perception. To address this, we propose F

Finite Certificates for In-Context Determinacy and a Threshold Theory of Emergence in Language Models

Model ReleasesDGX agent

arXiv:2606.07623v1 Announce Type: new Abstract: This paper develops a model-theoretic framework for verifying context-conditioned language-model behavior by replacing benchmark labels with finite sema

FIT-Print: Towards False-claim-resistant Model Ownership Verification via Targeted Fingerprint

Model ReleasesDGX agent

arXiv:2501.15509v5 Announce Type: replace-cross Abstract: Model fingerprinting has emerged as a crucial mechanism for safeguarding the intellectual property of open-source models, offering a non-intru

FlashMemory-DeepSeek-V4: Lightning Index Ultra-Long Context via Lookahead Sparse Attention

Model ReleasesDGX agent

arXiv:2606.09079v1 Announce Type: cross Abstract: Conventional LLMs keep the full KV cache loaded during decoding, causing a severe GPU memory bottleneck for ultra-long context serving. In this report

Fluid, natural voice translation with Gemini 3.5 Live Translate

Model ReleasesDGX agent

Gemini 3.5 Live Translate is an audio model delivering near real-time speech-to-speech translation in over 70 languages , with automatic language detection and natural-sounding translated speech that

FMRFusion: Frequency-Aware Multi-View Representation Learning for Heterogeneous Image Fusion

Model ReleasesDGX agent

arXiv:2606.07985v1 Announce Type: new Abstract: Infrared and visible image fusion aims to generate a composite image that retains significant target information and preserves detailed textures, integr

Fourier fractal dimension to predict the generalization of deep neural networks

Model ReleasesDGX agent

arXiv:2606.08308v1 Announce Type: new Abstract: Predicting the generalization performance of deep neural networks without relying on hold-out validation data is a fundamental challenge in machine lear

Frequency-Domain Latent Attention Gating for Cross-Domain Token Aggregation

Model ReleasesDGX agent

arXiv:2606.08191v1 Announce Type: cross Abstract: Token aggregation is a common bottleneck in models that map token representations to sample-level predictions, yet most pooling methods operate only i

From Coarse to Fine: Managing Temporal Granularity in Spatio-Temporal Data for Fine-Grained Traffic Prediction

Model ReleasesDGX agent

arXiv:2606.09392v1 Announce Type: new Abstract: Efficient acquisition, storage, and utilization of traffic data are critical challenges in spatio-temporal data management. Most traffic data systems co

From Hazard Functions to Language Space: Cox-Supervised Distillation of Survival Risk into a Large Language Model

Model ReleasesDGX agent

arXiv:2606.08945v1 Announce Type: new Abstract: We investigate whether information about time-to-event risk estimated by a Cox proportional hazards model can be transferred into a generative large lan

From Human Guidance to Autonomy: Agent Skill System for End-to-End LLM Deployment on Spatial NPUs

Model ReleasesDGX agent

arXiv:2606.07586v1 Announce Type: cross Abstract: Spatial neural processing units (NPUs) provide an energy-efficient platform for edge LLM inference, but efficiently deploying an LLM end-to-end on suc

From `May' to `Is': Certainty Distortion in Language Model Rewriting

Model ReleasesDGX agent

arXiv:2606.07951v1 Announce Type: cross Abstract: Humans increasingly turn to Language Models (LMs) in ways that shape beliefs and drive decisions, including discussing, rewriting, and summarizing inf

From Rigid to Dynamic: Entropy-Guided Adaptive Inference for Long-Context LLMs

Model ReleasesDGX agent

arXiv:2606.09508v1 Announce Type: new Abstract: Existing sparse attention and KV cache compression methods for long-context LLM inference typically apply fixed sparsity patterns or uniform budgets acr

From Statute to Control Flow: Span-Grounded Deontic Trees for Defeasible Scope Parsing

Model ReleasesDGX agent

arXiv:2606.08932v1 Announce Type: cross Abstract: Rule-following agents tasked with executing policies and regulations often fail via Silent Scope Omission (SSO): a model applies a general rule but si

FunctionEvolve: Structure-Guided Symbolic Regression with LLMs

Model ReleasesDGX agent

arXiv:2606.07704v1 Announce Type: cross Abstract: Symbolic regression aims to uncover explicit scientific laws from data. Recent methods use LLMs to guide mutation from background text, which is more

GD-MIL: Grade-Disentangled Multiple Instance Learning for Multimodal Biochemical Recurrence Prediction in Prostate Cancer

Model ReleasesDGX agent

arXiv:2606.09453v1 Announce Type: new Abstract: Biochemical recurrence (BCR) after radical prostatectomy is a critical endpoint in prostate cancer, yet risk stratification relies almost entirely on va

GEAR-VLA: Learning Geometry-Aware Action Representations for Generalizable Robotic Manipulation

Model ReleasesDGX agent

arXiv:2606.08530v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models achieve strong benchmark performance but still struggle in real-world deployment with unseen objects, background s

Gemini for Government: Your blueprint for mission impact

Model ReleasesDGX agent

The public sector has reached a critical inflection point. For years, organizations have explored what’s possible through isolated AI pilots and experimentation. Today, the question has shifted to “wh

Generalization Error Curves for Analytic Spectral Algorithms under Power-law Decay

Model ReleasesDGX agent

arXiv:2401.01599v4 Announce Type: replace Abstract: The generalization error curve of certain kernel regression method aims at determining the exact order of generalization error with various source c

Generalization in Nonlinear Least Squares via Learned Feature Geometry

Model ReleasesDGX agent

arXiv:2606.08799v1 Announce Type: cross Abstract: We study the generalization of ridge-regularized nonlinear least-squares models via on-average algorithmic stability, deriving error bounds for local

GIScholarBench: Benchmarking LLM Overconfidence in GIS Research

Model ReleasesDGX agent

arXiv:2606.08036v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in academic research workflows, but scholarly tasks require high factual precision and therefore ex

GlobeAudio: A Multilingual Multicultural Benchmark for Naturalistic Evaluation of Large Audio-Language Models

Model ReleasesDGX agent

arXiv:2606.08194v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) integrate audio perception and language understanding within a unified framework, enabling a wide range of real-wo

Google announces Gemini 3.5 Live Translate for instant voice-to-voice translation

Model ReleasesDGX agent

Gemini 3.5 Live Translate is Google's latest audio model delivering near real-time speech-to-speech translation in over 70 languages. The model automatically detects 70+ languages and generates smooth

Google upgrades NotebookLM to Gemini 3.5, adds more coding features

Model ReleasesDGX agent

Google LLC today updated its NotebookLM service with a set of online research and coding features designed to save time for users. NotebookLM is part note-taking app, part data analysis tool. Workers

Graph2Idea:Retrieval-Augmented Scientific Idea Generation with Graph-Structured Contexts

Model ReleasesDGX agent

arXiv:2606.09105v1 Announce Type: new Abstract: Generating novel, feasible, and high-quality research ideas is an important yet challenging task in scientific discovery.Recent Large Language Model (LL

GraphLoRA: Structure-Aware Low-Rank Adaptation for Large Language Model Recommendation

Model ReleasesDGX agent

arXiv:2606.07526v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown strong potential for recommendation (LLMRec) due to their powerful reasoning and generalization abilities. How

GRPO Does Not Close the Multi-Agent Coordination Gap

Model ReleasesDGX agent

arXiv:2606.07845v1 Announce Type: cross Abstract: We measure how well current large language models coordinate as multiple agents sharing a common resource, using the dining philosophers problem as a

HA-VLN 2.0: An Open Benchmark and Leaderboard for Human-Aware Navigation in Discrete and Continuous Environments with Dynamic Multi-Human Interactions

Model ReleasesDGX agent

arXiv:2503.14229v4 Announce Type: replace Abstract: Vision-and-Language Navigation (VLN) has been studied mainly in either discrete or continuous spaces, with little attention to dynamic, crowded envi

← Previous
1…156157158159160…377
Next →