AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Pointer-CAD v2: Plan-Then-Construct CAD Generation with Dimension-Aware Parametric Precision

DGX agent

arXiv:2606.29301v1 Announce Type: new Abstract: Computer-aided design (CAD) plays a fundamental role in modern manufacturing by providing the high precision required for industrial production. Recent

model-releasesarxiv-cs-cv
30 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

PolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM Agents

DGX agent

arXiv:2606.29225v1 Announce Type: new Abstract: LLM agents handle user requests on behalf of organizations through tool calls and must follow the company policies stated in their system prompts. Prior

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Pooled Leaderboards Hide System-Specific Winners: A Reporting-Protocol Audit of Offline Root-Cause Analysis Benchmarks

DGX agent

arXiv:2606.29159v1 Announce Type: new Abstract: Offline root-cause-analysis (RCA) benchmarks commonly rank methods by a single pooled top-1 accuracy across multiple subsystems, and engineers often rea

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

PoseShield: Neural Collision Fields for Human Self-Collision Resolution

DGX agent

arXiv:2606.29686v1 Announce Type: new Abstract: Self-collision remains a persistent challenge in SMPL-based human pose estimation and motion generation. Under extreme articulations or stochastic motio

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Position: RL Researchers Need to Distinguish Between Solving Simulators and Using Simulators as a Proxy

DGX agent

arXiv:2606.28433v1 Announce Type: new Abstract: One goal in reinforcement learning (RL) research is to understand general-purpose sequential decision-making, using benchmark simulators as a proxy for

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Post-training for Efficient Communication via Convention Formation

DGX agent

arXiv:2508.06482v2 Announce Type: replace-cross Abstract: Humans communicate with increasing efficiency in multi-turn interactions, by adapting their language and forming ad-hoc conventions. In contra

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Primary ICD Category Prediction using LLM-based Probing

DGX agent

arXiv:2606.28798v1 Announce Type: new Abstract: Objective: ICD codes are central to reimbursement, research, and population health surveillance, yet automated coding systems often struggle to integrat

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Probabilistic Approach to Black-Box Binary Optimization with Budget Constraints: Application to Sensor Placement

DGX agent

arXiv:2406.05830v2 Announce Type: replace-cross Abstract: This paper presents a fully probabilistic approach for solving optimal experimental design problems under budget constraints. The experimental

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Progressive Self-Supervised Learning with Individualized Community Assignment for Brain Network Analysis

DGX agent

arXiv:2606.29695v1 Announce Type: new Abstract: Brain networks exhibit a modular community structure that varies across individuals and neurological conditions. However, existing self-supervised learn

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Projected Exploitability Descent for Nash Equilibrium Computation in Multiplayer Imperfect-Information Games

DGX agent

arXiv:2606.29169v1 Announce Type: cross Abstract: Many important games have more than two players and imperfect information. Existing approaches for computing Nash equilibrium, the central game-theore

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Proteus: Automated Adversarial Robustness Testing for Audio Deepfake Detectors

DGX agent

arXiv:2606.29544v1 Announce Type: cross Abstract: We present Proteus, a framework developed at Resemble AI for automated robustness testing of our audio deepfake detection system. Given a detector, Pr

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization

DGX agent

arXiv:2602.11351v2 Announce Type: replace Abstract: Proactive large language model (LLM) agents aim to actively plan, query, and interact over multiple turns, enabling efficient task completion beyond

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System

DGX agent

arXiv:2606.18112v3 Announce Type: replace-cross Abstract: Agentic navigation systems require a base navigation model whose observation strategy can be externally reconfigured at inference time, becaus

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

RA-QA: A Benchmarking System for Respiratory Audio Question Answering Under Real-World Heterogeneity

DGX agent

arXiv:2602.18452v3 Announce Type: replace-cross Abstract: As conversational multimodal AI tools are increasingly adopted to process patient data for health assessment, robust benchmarks are needed to

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Randomized neural operator for parametric PDEs with fast training and conformal uncertainty quantification

DGX agent

arXiv:2606.29440v1 Announce Type: new Abstract: Repeatedly solving parametric PDEs is essential for uncertainty quantification, design optimization and inverse problems, but conventional neural operat

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

RankGraph-2: Lifecycle Co-Design for Billion-Node Graph Learning in Recommendation

DGX agent

arXiv:2606.18379v2 Announce Type: replace-cross Abstract: Graph-based retrieval at billion-node scale requires jointly solving three tightly coupled problems -- graph construction, representation lear

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Reachability Guarantees for Cart-Pole Swing-Up and Stabilization

DGX agent

arXiv:2606.28627v1 Announce Type: cross Abstract: The cart-pole swing-up is a canonical benchmark for nonlinear control of underactuated systems, yet an end-to-end guarantee linking the global swing-u

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

Recursive Self-Evolving Agents via Held-Out Selection

DGX agent

arXiv:2606.28374v1 Announce Type: new Abstract: LLM agents are increasingly improved without weight updates by evolving a natural-language artifact, such as reflections, workflows, playbooks, cheatshe

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Redefining Maritime Anomaly Detection via Equation-Grounded Synthetic Anomalies

DGX agent

arXiv:2606.29721v1 Announce Type: cross Abstract: Maritime anomaly detection is essential for ensuring maritime safety, security, and efficient traffic management at sea, with Automatic Identification

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

RefAlign: Representation Alignment for Reference-to-Video Generation

DGX agent

arXiv:2603.25743v2 Announce Type: replace Abstract: Reference-to-video (R2V) generation is a controllable video synthesis paradigm that constrains the generation process using both text prompts and re

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Rehearsed Multi-Agent Live Product Demonstrations with Real-Time Voice Question Answering

DGX agent

arXiv:2606.30294v1 Announce Type: new Abstract: Live product demonstrations are a recurring, high-cost activity in software organizations: a human presenter must select features, dispatch the correspo

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Reliability-Prioritized Fine-Grained Generation in Multimodal Large

DGX agent

arXiv:2606.29573v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are increasingly expected to generate fine-grained descriptions of visual content. However, we observe and theo

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

REPAIR-Bench: A Benchmark for Robot Error Perception And Interaction Recovery

DGX agent

arXiv:2606.29937v1 Announce Type: new Abstract: Understanding how users perceive and respond to robot failures is essential for building robust and trustworthy robot systems. Prior work, however, (i)

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

Reported Confidence in LLMs Tracks Commitment More Than Correctness

DGX agent

arXiv:2606.29490v1 Announce Type: cross Abstract: Confidence is an estimate of the probability that a chosen answer is correct. Verbal confidence reports are widely used as uncertainty measures in lar

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models

DGX agent

arXiv:2606.29196v1 Announce Type: cross Abstract: Do language models know when they are being tested? This question matters for AI safety: a model that recognises an evaluation context could alter its

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Research Entity Extraction and Topic Detection from UKRI Grant Proposals

DGX agent

arXiv:2606.30304v1 Announce Type: cross Abstract: This paper presents preliminary findings from a UKRI-funded Metascience project comparing three LLM-based approaches, GPT-4o, Mistral, and a bespoke a

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Residual-Guided Dictionary Learning for Spectrally Accurate Koopman Approximation

DGX agent

arXiv:2606.29083v1 Announce Type: cross Abstract: Koopman theory promises linear structure in nonlinear dynamics, but numerical Koopman spectra are easy to compute and hard to trust. A finite EDMD mat

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Rethinking Generative Reconstruction Attacks against Graph Neural Network Models

DGX agent

arXiv:2606.29748v1 Announce Type: new Abstract: The application of graph data in numerous disciplines raises the need for gathering and analyzing huge volumes of data, some of which is private and sen

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs for Code Generation

DGX agent

arXiv:2606.28998v1 Announce Type: cross Abstract: Large Language Model (LLM) alignment trains an LLM using preference data to produce outputs that better meet established quality standards. While LLM

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

RIPA: Sensory-Vector Prompt Injection Attacks on LLM-Controlled ROS 2 Robots

DGX agent

arXiv:2606.28649v1 Announce Type: cross Abstract: We present RIPA, the first systematic multi-channel empirical study of prompt injection attacks delivered through the sensory pipeline of a ROS 2-base

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

RiverONE: Generating Knowledge-Intensive VLM by Simulated Quantum Machines

DGX agent

arXiv:2606.29966v1 Announce Type: cross Abstract: Quantum computing provides a powerful paradigm for representing and transforming high-dimensional information through superposition, entanglement, and

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

RoAd-RL: A Unified Library and Benchmark for Robust Adversarial Reinforcement Learning

DGX agent

arXiv:2606.29867v1 Announce Type: cross Abstract: Deep Reinforcement Learning (DRL) has achieved significant success in robotics and autonomous systems, yet remains vulnerable to adversarial perturbat

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

RoboGaze: Evaluating Robot World Models via Structured Vision-Language Analysis

DGX agent

arXiv:2606.28385v1 Announce Type: cross Abstract: Recent advances in robot world models enable synthetic video generation for embodied prediction and planning. However, evaluating these videos is chal

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

RSGPNet: Geometric Prompting for Remote Sensing Open-Vocabulary Semantic Segmentation

DGX agent

arXiv:2606.28410v1 Announce Type: cross Abstract: Open-vocabulary semantic segmentation (OVSS) enables text-guided segmentation of unseen objects, breaking fixed-class limitations to achieve open-worl

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intelligence

DGX agent

arXiv:2606.20515v2 Announce Type: replace Abstract: Real-world spatial intelligence requires reasoning over a continuous and evolving 3D world, yet existing VLMs and tool-augmented agents largely rema

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

SA-Homo: Scale Adaptive Homography Estimation for Scale Variation Scenarios

DGX agent

arXiv:2606.30408v1 Announce Type: new Abstract: Homography estimation, as one of the fundamental problems in computer vision, remains challenged by scale variation scenarios where image pairs potentia

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

SABER-Math: Automated Benchmark for Information Retrieval Evaluation in Mathematics

DGX agent

arXiv:2606.29894v1 Announce Type: cross Abstract: As agentic AI systems tackle more complex mathematical tasks, they increasingly rely on information retrieval (IR) to search problem databases, theore

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SADL: What to Ignore? A Benchmark for Subject-Aware Distractor Localization

DGX agent

arXiv:2606.30393v1 Announce Type: new Abstract: Photographs frequently contain visual distractors besides foregrounds and backgrounds of the intended subject, competing for attention and weakening com

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

SafePyramid: A Hierarchical Benchmark for In-context Policy Guardrailing

DGX agent

arXiv:2606.29887v1 Announce Type: new Abstract: In real-world applications, guardrails are often expected to identify unsafe user-model interactions according to application-specific safety policies,

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SAKE: Software Architectural Knowledge Evaluation Benchmark for Large Language Models

DGX agent

arXiv:2606.29520v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used as assistants across the software development lifecycle, yet their ability to reason about software

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SatSplat: Geometrically-Accurate Gaussian Splatting for Satellite Imagery

DGX agent

arXiv:2606.28581v1 Announce Type: new Abstract: High-resolution satellite imagery demands 3D reconstruction methods that deliver both speed and geometric accuracy. Recent adaptations of 3D Gaussian Sp

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Scalar Representations of Neural Network Training Dynamics

DGX agent

arXiv:2606.30384v1 Announce Type: new Abstract: Training in artificial neural networks can be viewed as a trajectory evolving through a high-dimensional loss landscape. However, the large number of tr

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

ScAle: Attention Head Scaling as a Minimal Adapter for Spatial Reasoning in Vision Language Models

DGX agent

arXiv:2606.29579v1 Announce Type: cross Abstract: Spatial reasoning remains a persistent challenge for many vision language models (VLMs), and improving it typically requires fine-tuning with substant

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Scaling Textual Gradients via Sampling-Based Momentum

DGX agent

arXiv:2506.00400v4 Announce Type: replace-cross Abstract: LLM-based prompt optimization, which uses LLM-provided ``textual gradients'' (feedback) to refine prompts, has emerged as an effective method

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Scaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B Agent

DGX agent

arXiv:2606.30616v1 Announce Type: new Abstract: We introduce Agents-A1, a 35B Mixture-of-Experts Agentic Model that reaches trillion-parameter-level performance by scaling the agent horizon. We invest

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

SCARCE: Scalable Cascade Analysis for Rare-event Characterisation via Embeddings

DGX agent

arXiv:2606.29623v1 Announce Type: new Abstract: Rare events govern the safety profile of modern AI systems, yet their probabilities are extremely difficult to estimate: direct Monte Carlo requires pro

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SciIR: A Large-scale Training Dataset and Benchmark for Scientific Image Reasoning Generation

DGX agent

arXiv:2606.30124v1 Announce Type: new Abstract: While Text-to-Image (T2I) models have shown remarkable success in generating photorealistic visual content, they still struggle with the rigorous semant

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

SciVisAgentBench: A Benchmark for Evaluating Scientific Data Analysis and Visualization Agents

DGX agent

arXiv:2603.29139v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have enabled agentic systems to translate natural-language intent into executable scientific visuali

model-releasesarxiv-cs-ai
30 Jun 2026
← Previous
1…108109110111112…361
Next →