AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,628 results
Model Releases

OptiMUS-0.3: Using Large Language Models to Model and Solve Optimization Problems at Scale

DGX agent

arXiv:2407.19633v4 Announce Type: replace Abstract: Optimization problems are pervasive in sectors from manufacturing and distribution to healthcare. However, most such problems are still solved heuri

model-releasesarxiv-cs-ai
30 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

ORCA: Open-ended Response Correctness Assessment for Audio Question Answering

DGX agent

arXiv:2512.09066v2 Announce Type: replace-cross Abstract: Reliable assessment of the abilities of large audio language models (LALMs) is essential to advancing the state of the art. As benchmarks rapi

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Ornith-1.0-35B is now available in claude code through hf-claude

DGX agent

Ornith-1.0-35B, a 35-billion parameter model, has been made available for use through Claude Code via Hugging Face integration. This announcement indicates expanded model availability and integration

model-releasesclem-delangue--x
30 Jun 2026
Model Releases

OSWorld2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks

DGX agent

arXiv:2606.29537v1 Announce Type: new Abstract: Existing computer-use benchmarks fail to capture the realism, complexity, and long-horizon demands of real-world computer use, limiting their ability to

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Parametric Skills

DGX agent

arXiv:2606.30015v1 Announce Type: new Abstract: Since intelligence fundamentally relies on efficient skill acquisition (Chollet, 2019), the ability to leverage skills is critical. For LLMs, skills, ma

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

PCGD: Physics-Guided Conditional Graph Diffusion for TCAD Device Simulation

DGX agent

arXiv:2606.29272v1 Announce Type: new Abstract: Technology computer-aided design (TCAD) semiconductor device simulation is fundamentally constrained by the high computational cost of iteratively solvi

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Perforce launches Agentic Gateway to govern AI agents and cut token costs

DGX agent

Perforce Software Inc. today expanded its Perforce Intelligence lineup with an agentic gateway for managing artificial intelligence agents, an autonomous testing platform driven by natural language an

model-releasessiliconangle
30 Jun 2026
Model Releases

PGE-SAM: Prompt-Guided Feature Enhancement for Interactive Segmentation under Degradation

DGX agent

arXiv:2606.30477v1 Announce Type: new Abstract: Segment Anything Model (SAM) has revolutionized promptable image segmentation with strong zero-shot generalization. However, its performance degrades su

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Pie launches with $19.5M to bring AI marketing to small businesses

DGX agent

Pie Tech Inc., a startup using artificial intelligence to provide growth tools for small businesses, today officially launched with an announcement that it has raised 19.5 million in new funding to ex

model-releasessiliconangle
30 Jun 2026
Model Releases

PlantExpertVQA: A Visual Question Answering Dataset for Benchmarking Vision-Language Models in Plant Science

DGX agent

arXiv:2508.17117v3 Announce Type: replace-cross Abstract: Existing plant-disease datasets target classification and detection, leaving vision-language models unable to support interactive, reasoning-b

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Pointer-CAD v2: Plan-Then-Construct CAD Generation with Dimension-Aware Parametric Precision

DGX agent

arXiv:2606.29301v1 Announce Type: new Abstract: Computer-aided design (CAD) plays a fundamental role in modern manufacturing by providing the high precision required for industrial production. Recent

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

PolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM Agents

DGX agent

arXiv:2606.29225v1 Announce Type: new Abstract: LLM agents handle user requests on behalf of organizations through tool calls and must follow the company policies stated in their system prompts. Prior

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Pooled Leaderboards Hide System-Specific Winners: A Reporting-Protocol Audit of Offline Root-Cause Analysis Benchmarks

DGX agent

arXiv:2606.29159v1 Announce Type: new Abstract: Offline root-cause-analysis (RCA) benchmarks commonly rank methods by a single pooled top-1 accuracy across multiple subsystems, and engineers often rea

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

PoseShield: Neural Collision Fields for Human Self-Collision Resolution

DGX agent

arXiv:2606.29686v1 Announce Type: new Abstract: Self-collision remains a persistent challenge in SMPL-based human pose estimation and motion generation. Under extreme articulations or stochastic motio

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Position: RL Researchers Need to Distinguish Between Solving Simulators and Using Simulators as a Proxy

DGX agent

arXiv:2606.28433v1 Announce Type: new Abstract: One goal in reinforcement learning (RL) research is to understand general-purpose sequential decision-making, using benchmark simulators as a proxy for

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Post-training for Efficient Communication via Convention Formation

DGX agent

arXiv:2508.06482v2 Announce Type: replace-cross Abstract: Humans communicate with increasing efficiency in multi-turn interactions, by adapting their language and forming ad-hoc conventions. In contra

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Primary ICD Category Prediction using LLM-based Probing

DGX agent

arXiv:2606.28798v1 Announce Type: new Abstract: Objective: ICD codes are central to reimbursement, research, and population health surveillance, yet automated coding systems often struggle to integrat

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Probabilistic Approach to Black-Box Binary Optimization with Budget Constraints: Application to Sensor Placement

DGX agent

arXiv:2406.05830v2 Announce Type: replace-cross Abstract: This paper presents a fully probabilistic approach for solving optimal experimental design problems under budget constraints. The experimental

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Progressive Self-Supervised Learning with Individualized Community Assignment for Brain Network Analysis

DGX agent

arXiv:2606.29695v1 Announce Type: new Abstract: Brain networks exhibit a modular community structure that varies across individuals and neurological conditions. However, existing self-supervised learn

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Projected Exploitability Descent for Nash Equilibrium Computation in Multiplayer Imperfect-Information Games

DGX agent

arXiv:2606.29169v1 Announce Type: cross Abstract: Many important games have more than two players and imperfect information. Existing approaches for computing Nash equilibrium, the central game-theore

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Proteus: Automated Adversarial Robustness Testing for Audio Deepfake Detectors

DGX agent

arXiv:2606.29544v1 Announce Type: cross Abstract: We present Proteus, a framework developed at Resemble AI for automated robustness testing of our audio deepfake detection system. Given a detector, Pr

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization

DGX agent

arXiv:2602.11351v2 Announce Type: replace Abstract: Proactive large language model (LLM) agents aim to actively plan, query, and interact over multiple turns, enabling efficient task completion beyond

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Qwen publishes new work on RL coding agents. (bookmark it) The idea is to continually build a verification system that co-evolves with AI ag…

DGX agent

Qwen publishes new work on RL coding agents. (bookmark it) The idea is to continually build a verification system that co-evolves with AI agents. LLMs suffer from all sorts of reward hacking issues. T

model-releasesdair-ai--x
30 Jun 2026
Model Releases

Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System

DGX agent

arXiv:2606.18112v3 Announce Type: replace-cross Abstract: Agentic navigation systems require a base navigation model whose observation strategy can be externally reconfigured at inference time, becaus

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

RA-QA: A Benchmarking System for Respiratory Audio Question Answering Under Real-World Heterogeneity

DGX agent

arXiv:2602.18452v3 Announce Type: replace-cross Abstract: As conversational multimodal AI tools are increasingly adopted to process patient data for health assessment, robust benchmarks are needed to

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Randomized neural operator for parametric PDEs with fast training and conformal uncertainty quantification

DGX agent

arXiv:2606.29440v1 Announce Type: new Abstract: Repeatedly solving parametric PDEs is essential for uncertainty quantification, design optimization and inverse problems, but conventional neural operat

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

RankGraph-2: Lifecycle Co-Design for Billion-Node Graph Learning in Recommendation

DGX agent

arXiv:2606.18379v2 Announce Type: replace-cross Abstract: Graph-based retrieval at billion-node scale requires jointly solving three tightly coupled problems -- graph construction, representation lear

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Reachability Guarantees for Cart-Pole Swing-Up and Stabilization

DGX agent

arXiv:2606.28627v1 Announce Type: cross Abstract: The cart-pole swing-up is a canonical benchmark for nonlinear control of underactuated systems, yet an end-to-end guarantee linking the global swing-u

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

Read the full Aston Martin F1 team interview with @aidangomez here: https://www.astonmartinf1.com/en-GB/news/feature/perspectives-aidan-gome…

DGX agent

This post links to a full interview with Aidan Gomez conducted by the Aston Martin F1 team, published on their official website under their 'Perspectives' feature section. The interview likely covers

model-releasescohere--x
30 Jun 2026
Model Releases

Recursive Self-Evolving Agents via Held-Out Selection

DGX agent

arXiv:2606.28374v1 Announce Type: new Abstract: LLM agents are increasingly improved without weight updates by evolving a natural-language artifact, such as reflections, workflows, playbooks, cheatshe

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Redefining Maritime Anomaly Detection via Equation-Grounded Synthetic Anomalies

DGX agent

arXiv:2606.29721v1 Announce Type: cross Abstract: Maritime anomaly detection is essential for ensuring maritime safety, security, and efficient traffic management at sea, with Automatic Identification

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

RefAlign: Representation Alignment for Reference-to-Video Generation

DGX agent

arXiv:2603.25743v2 Announce Type: replace Abstract: Reference-to-video (R2V) generation is a controllable video synthesis paradigm that constrains the generation process using both text prompts and re

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Rehearsed Multi-Agent Live Product Demonstrations with Real-Time Voice Question Answering

DGX agent

arXiv:2606.30294v1 Announce Type: new Abstract: Live product demonstrations are a recurring, high-cost activity in software organizations: a human presenter must select features, dispatch the correspo

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Reliability-Prioritized Fine-Grained Generation in Multimodal Large

DGX agent

arXiv:2606.29573v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are increasingly expected to generate fine-grained descriptions of visual content. However, we observe and theo

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

REPAIR-Bench: A Benchmark for Robot Error Perception And Interaction Recovery

DGX agent

arXiv:2606.29937v1 Announce Type: new Abstract: Understanding how users perceive and respond to robot failures is essential for building robust and trustworthy robot systems. Prior work, however, (i)

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

Reported Confidence in LLMs Tracks Commitment More Than Correctness

DGX agent

arXiv:2606.29490v1 Announce Type: cross Abstract: Confidence is an estimate of the probability that a chosen answer is correct. Verbal confidence reports are widely used as uncertainty measures in lar

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models

DGX agent

arXiv:2606.29196v1 Announce Type: cross Abstract: Do language models know when they are being tested? This question matters for AI safety: a model that recognises an evaluation context could alter its

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Research Entity Extraction and Topic Detection from UKRI Grant Proposals

DGX agent

arXiv:2606.30304v1 Announce Type: cross Abstract: This paper presents preliminary findings from a UKRI-funded Metascience project comparing three LLM-based approaches, GPT-4o, Mistral, and a bespoke a

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Residual-Guided Dictionary Learning for Spectrally Accurate Koopman Approximation

DGX agent

arXiv:2606.29083v1 Announce Type: cross Abstract: Koopman theory promises linear structure in nonlinear dynamics, but numerical Koopman spectra are easy to compute and hard to trust. A finite EDMD mat

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Rethinking Generative Reconstruction Attacks against Graph Neural Network Models

DGX agent

arXiv:2606.29748v1 Announce Type: new Abstract: The application of graph data in numerous disciplines raises the need for gathering and analyzing huge volumes of data, some of which is private and sen

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs for Code Generation

DGX agent

arXiv:2606.28998v1 Announce Type: cross Abstract: Large Language Model (LLM) alignment trains an LLM using preference data to produce outputs that better meet established quality standards. While LLM

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

RIPA: Sensory-Vector Prompt Injection Attacks on LLM-Controlled ROS 2 Robots

DGX agent

arXiv:2606.28649v1 Announce Type: cross Abstract: We present RIPA, the first systematic multi-channel empirical study of prompt injection attacks delivered through the sensory pipeline of a ROS 2-base

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

RiverONE: Generating Knowledge-Intensive VLM by Simulated Quantum Machines

DGX agent

arXiv:2606.29966v1 Announce Type: cross Abstract: Quantum computing provides a powerful paradigm for representing and transforming high-dimensional information through superposition, entanglement, and

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

RoAd-RL: A Unified Library and Benchmark for Robust Adversarial Reinforcement Learning

DGX agent

arXiv:2606.29867v1 Announce Type: cross Abstract: Deep Reinforcement Learning (DRL) has achieved significant success in robotics and autonomous systems, yet remains vulnerable to adversarial perturbat

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

RoboGaze: Evaluating Robot World Models via Structured Vision-Language Analysis

DGX agent

arXiv:2606.28385v1 Announce Type: cross Abstract: Recent advances in robot world models enable synthetic video generation for embodied prediction and planning. However, evaluating these videos is chal

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Room 2016 for those attending @aiDotEngineer 2:25pm. Will also cover Galactica, early Llama reasoning efforts and more - think this is the f…

DGX agent

This post announces a conference session at @aiDotEngineer scheduled for 2:25pm in Room 2016, covering topics including the Galactica model and early reasoning efforts in Llama models, with additional

model-releasesswyx--x
30 Jun 2026
Model Releases

RSGPNet: Geometric Prompting for Remote Sensing Open-Vocabulary Semantic Segmentation

DGX agent

arXiv:2606.28410v1 Announce Type: cross Abstract: Open-vocabulary semantic segmentation (OVSS) enables text-guided segmentation of unseen objects, breaking fixed-class limitations to achieve open-worl

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intelligence

DGX agent

arXiv:2606.20515v2 Announce Type: replace Abstract: Real-world spatial intelligence requires reasoning over a continuous and evolving 3D world, yet existing VLMs and tool-augmented agents largely rema

model-releasesarxiv-cs-cv
30 Jun 2026
← Previous
1…153154155156157…472
Next →