AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,098 results
11 Aug 2026

OpenVisTool: An Open Recipe for Synthesizing Instructive Visual Tool-Use Trajectories

Model ReleasesDGX agent

arXiv:2608.08557v1 Announce Type: new Abstract: Visual tool use has emerged as a fundamental capability for multimodal agents to actively acquire evidence beyond a fixed image encoding. The prevailing

OpenWALDO launches to build collaborative community for open-source AI

Model ReleasesDGX agent

OpenWALDO, a new open-source artificial intelligence project sponsored by Ctrl IQ Inc., launched today, led by Gregory Kutzer, the founder of Rocky Linux, CentOS and Apptainer. The project aims to bui

Optimal Multi-Agent Path Finding in Continuous Time

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2508.16410v3 Announce Type: replace-cross Abstract: Continuous-time Conflict Based Search (CCBS) has been widely used as an exact baseline for Continuous-time Multi-Agent Path Finding (MAPFR), a

Other active promotions: - Free models: Solar Pro 4 (1 week), Hy3, Step 3.7 Flash, Laguna S and XS - 90% off DeepSeek V4 Flash for ~2 more d…

Model ReleasesDGX agent

Nous Research has extended its 20 % discount on all models—including high‑end frontier options—throughout the Nous Portal for an additional two weeks (until the end of April). Free model trials such a

Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution

Model ReleasesDGX agent

arXiv:2608.08311v1 Announce Type: cross Abstract: We present Ouroboros, a self-developing agent harness whose tools, prompts, context assembly, and core implementation improve through reviewed commits

Out-of-Distribution Federated Distillation with Domain-Aware Proxy

Model ReleasesDGX agent

arXiv:2608.08525v1 Announce Type: new Abstract: Federated Learning is a distributed machine learning paradigm that trains a global model by aggregating local clients without sharing private data of ea

P^{3}: Joint Program-and-Proof Planning for Verified Code Generation

Model ReleasesDGX agent

arXiv:2608.09277v1 Announce Type: new Abstract: Verified code generation asks a large language model (LLM) to generate both an executable program and a machine-checkable proof that the program meets a

PACE: A Playback-Aligned Context Engine for LLM-Based Full-Duplex Voice Dialogue

Model ReleasesDGX agent

arXiv:2608.07631v1 Announce Type: cross Abstract: LLM-based full-duplex voice services allow users to speak while the assistant is responding. Because servers can generate output and advance dialogue

Parameter-Dependent LMI Synthesis for Semi-Global Differential ISS Trajectory Tracking of Nonholonomic Mobile Robots Under Multiplicative Wheel Slip

Model ReleasesDGX agent

arXiv:2608.08049v1 Announce Type: cross Abstract: This paper presents a parameter-dependent linear matrix inequality (LMI) framework for trajectory tracking of nonholonomic mobile robots subject to se

Parameter Exploration for RLVR via Variational Learning

Model ReleasesDGX agent

arXiv:2608.09805v1 Announce Type: cross Abstract: Exploration has been a focus of reinforcement learning research for a long time. Recently, there has been growing evidence that it is also an importan

Path-dependent Discrete Amortized Inference

Model ReleasesDGX agent

arXiv:2608.08644v1 Announce Type: new Abstract: We consider the problem of sampling compositional and discrete objects from a given unnormalized posterior distribution. Notably, recent studies have sh

Performance of large language models in the optical diagnosis of colorectal polyps

Model ReleasesDGX agent

arXiv:2608.07543v1 Announce Type: cross Abstract: Background and Study Aims: Accurate optical diagnosis of colorectal polyps guides resection strategy and surveillance, with multimodal large language

Persistent Semantic Entities in Tool-Augmented LLM Systems

Model ReleasesDGX agent

arXiv:2608.07952v1 Announce Type: cross Abstract: Tool-augmented LLM agents can harbor implicit state that persists across sessions, activates through events, and propagates across agent boundaries---

Personalized Federated Learning via Variance-Aware Nonparametric Empirical Bayes

Model ReleasesDGX agent

arXiv:2608.09074v1 Announce Type: cross Abstract: We develop a new approach to Personalized Federated Learning across heterogeneous clients using Nonparametric Empirical Bayes (NPEB). Leveraging the a

Personalized Lower-limb Exoskeleton Assistance via Preference-based Bayesian Optimization

Model ReleasesDGX agent

arXiv:2608.09015v1 Announce Type: new Abstract: A significant challenge in exoskeleton robotics is the need to dynamically adapt control profiles to individual motion preferences, thereby ensuring bot

PluginEval: A Diagnostic Benchmark for Fine-Grained Error Attribution in Function Calling

Model ReleasesDGX agent

arXiv:2608.08700v1 Announce Type: new Abstract: Reliable evaluation of tool routing is critical as Large Language Models increasingly operate as autonomous agents. Current benchmarks face three struct

PolicyKG: An Agentic LLM Pipeline for Translating Institutional Policies into SHACL Knowledge Graphs

Model ReleasesDGX agent

arXiv:2608.09028v1 Announce Type: new Abstract: Institutional policies stay in natural language while the systems that check compliance demand machine-readable constraints. Bridging that gap is still

Population-Level Generative Modeling for Ranking Data

Model ReleasesDGX agent

arXiv:2608.08422v1 Announce Type: cross Abstract: Ranking data arise in scientific and machine learning applications, including recommendation systems, information retrieval, voting, marketing, and AI

PragMatch: Separating Pragmatic Incongruity from Cross-Modal Mismatch in Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2608.09772v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have demonstrated strong performance on multimodal benchmarks, yet it remains unclear whether they genuinely reason

Predicting blood clot growth from sparse post-onset measurements with latent neural differential equations

Model ReleasesDGX agent

arXiv:2608.08165v1 Announce Type: new Abstract: Computational models of blood clotting improve understanding of thrombus formation, but their clinical application remains limited because many model in

Predictive safety filter enhanced curriculum learning control for efficient vehicle dynamics controller

Model ReleasesDGX agent

arXiv:2608.09653v1 Announce Type: cross Abstract: Recent advances in learning-based control have enabled impressive achievements in solving complex control problems in various domains. However, since

Preserving Item Semantics for Free: Rethinking Token Initialization in LLM-Based Generative Recommendation

Model ReleasesDGX agent

arXiv:2608.07816v1 Announce Type: cross Abstract: Recent advances in generative recommendation (GR) leverage large language models (LLMs) as recommender backbones, enabling LLMs to directly generate r

Probabilistic Circuits for Knowledge Graph Completion with Reduced Rule Sets

Model ReleasesDGX agent

arXiv:2508.06706v2 Announce Type: replace Abstract: Rule-based methods for knowledge graph completion provide explainable results, but often require tens of thousands of rules to achieve competitive p

Prompt engineering does not universally improve Large Language Model performance across clinical decision-making tasks

Model ReleasesDGX agent

arXiv:2512.22966v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated promise in medical knowledge assessments, yet their practical utility in real-world clinical decision

PROSLEX: A Novel Dataset for Expert-Annotated Legal Statute Prediction for Indian Judiciary

Model ReleasesDGX agent

arXiv:2608.08830v1 Announce Type: new Abstract: Legal Statute Prediction (LSP) involves automatically identifying relevant legal statutes given factual descriptions in legal documents, typically frame

Psychological methods really work. Has anyone tried to encouraging and instilling confidence to GPT?

Model ReleasesDGX agent

Oddly enough, encouragement actually influences the performance of not only Claude but also GPT and other AIs. While Claude was working on a complex problem related to the Riemann Hypothesis, the Anth

QuArch: A Benchmark for Evaluating LLM Reasoning in Computer Architecture

Model ReleasesDGX agent

arXiv:2510.22087v3 Announce Type: replace-cross Abstract: The field of computer architecture, which bridges high-level software abstractions and low-level hardware implementations, remains absent from

Quokka: Accelerating Program Verification with LLMs via Invariant Synthesis

Model ReleasesDGX agent

arXiv:2509.21629v4 Announce Type: replace-cross Abstract: Program verification relies on loop invariants, yet automatically discovering strong invariants remains a long-standing challenge. We investig

RA-FinBERT: Rule-aware LoRA adaptation for low-resource financial sentiment classification

Model ReleasesDGX agent

arXiv:2608.09834v1 Announce Type: new Abstract: Financial sentiment analysis converts unstructured financial news into quantitative signals that can support market analysis and decision-making. Existi

RAG-Based Auto-Configuration for Industrial Fieldbus Devices

Model ReleasesDGX agent

arXiv:2608.08618v1 Announce Type: cross Abstract: Industrial device commissioning requires engineers to manually extract hundreds of protocol-specific parameters from heterogeneous PDF manuals and tra

RAGMesh with FaME-G2E: Long-Form Text-Driven 3D Face Generation and Editing

Model ReleasesDGX agent

arXiv:2608.09186v1 Announce Type: new Abstract: Text-driven 3D face generation and editing remains challenging due to the difficulty of translating long-form descriptions into fine-grained facial geom

Readout-Rank Laws for Isotropic Quantum Tangents

Model ReleasesDGX agent

arXiv:2608.07628v1 Announce Type: cross Abstract: Deep parameterized quantum circuits may remain sensitive to a parameter change while the observables retained by a learning model barely respond. We s

Real Data Closes Synthetic-to-Real Gap in Optical Chemical Structure Recognition

Model ReleasesDGX agent

arXiv:2608.09100v1 Announce Type: cross Abstract: Millions of chemical structures appear in patents and papers only as drawings, and using that information at scale requires reading the drawings. OCSR

RealDenseFace: Real-time Monocular 3D Face Reconstruction from Dense UV-space Priors

Model ReleasesDGX agent

arXiv:2608.09238v1 Announce Type: new Abstract: Recent monocular 3D face reconstruction methods achieve high fidelity by fitting a 3D Morphable Model (3DMM) to dense priors predicted by networks, but

really excited to see this release AND integrate it with deepagents!! https://www.langchain.com/blog/switchyard-agent-routing-benchmark

Model ReleasesDGX agent

really excited to see this release AND integrate it with deepagents!! https://www.langchain.com/blog/switchyard-agent-routing-benchmark Lightning strikes for continuous and long-run agents! Nemotron 3

Reason Wide, Not Deep: Amortizing the Reasoning Premium into Distilled Skills

Model ReleasesDGX agent

arXiv:2608.07885v1 Announce Type: new Abstract: Reasoning modes of language models outperform their non-reasoning counterparts on multi-step agentic tasks, but pay a 3-6x premium in output tokens on e

RecoverFly: A Failure-Aware Reinforcement Learning Post-Training Framework for Aerial Vision-Language Navigation

Model ReleasesDGX agent

arXiv:2608.09467v1 Announce Type: cross Abstract: Unmanned aerial vehicle vision-language navigation (UAV-VLN) requires agents to translate visual observations and language instructions into reliable

Reducing Pretraining-Generation Mismatch in Diffusion Language Models

Model ReleasesDGX agent

arXiv:2608.09424v1 Announce Type: new Abstract: Autoregressive language models align training and use: generation conditions on a clean prompt, and training predicts future tokens from clean left cont

REFRAMED: Towards Realistic Audio Description Generation for Movies

Model ReleasesDGX agent

arXiv:2608.09765v1 Announce Type: new Abstract: Audio Description (AD) is a verbal narration of key visual content in videos, enabling access for visually impaired audiences. Unlike standard video cap

ReliableNet: A Chance-Constrained Approach to Trustworthy Classification in Deep Learning

Model ReleasesDGX agent

arXiv:2608.09768v1 Announce Type: new Abstract: A prediction that is both confident and wrong is a critical reliability failure because it can bypass abstention and human review precisely when the mod

REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation

Model ReleasesDGX agent

arXiv:2503.22122v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in robotic planning, particularly for long-horizon tasks that require

RenderMatte: Exact-Alpha Rendering and Group-Relative Alignment for Image Matting

Model ReleasesDGX agent

arXiv:2608.08487v1 Announce Type: new Abstract: Image matting is an essential enabling technology for modern visual content production, where foreground extraction determines the realism and editabili

Reproducing and Stress-Testing Two Approaches to LLM Reasoning Reliability: Test-Time Probability Aggregation and Logic-Representation Editing

Model ReleasesDGX agent

arXiv:2608.08514v1 Announce Type: new Abstract: We independently reproduce two recent methods for making large language model (LLM) reasoning more reliable, and stress-test them across domains and mod

Researchers find that feeding a frontier model's encrypted reasoning traces to a weaker model from the same provider can make it output the traces in plaintext (Will Knight/Wired)

Model ReleasesDGX agent

Will Knight / Wired: Researchers find that feeding a frontier model's encrypted reasoning traces to a weaker model from the same provider can make it output the traces in plaintext — Researchers devis

Rethinking 3D Segmentation from Individual LiDAR Scans: Incidence-Aware Sampling on the SIP Benchmark

Model ReleasesDGX agent

arXiv:2608.07757v1 Announce Type: new Abstract: 3D scene understanding is increasingly important in construction, yet most methods are developed on curated datasets that do not fully reflect real site

Rethinking Attention Locality in Spiking Transformers

Model ReleasesDGX agent

arXiv:2608.08541v1 Announce Type: new Abstract: Spiking Transformers provide a promising paradigm for efficient visual processing with spike-driven computation, yet their Softmax-free Spiking Self-Att

Rethinking Medical Landmark Localization with Prototype Learning-based Progressive Offset Correction

Model ReleasesDGX agent

arXiv:2608.09182v1 Announce Type: cross Abstract: Accurate landmark localization in medical images is a fundamental step for quantitative clinical measurement and downstream analysis. Existing localiz

Rethinking Self-Evolving Agents: Do We Still Need Prescribed Optimization Pipelines?

Model ReleasesDGX agent

arXiv:2608.09629v1 Announce Type: new Abstract: Self-evolving agents are usually built around prescribed optimization pipelines: the framework decides how to gather evidence, revise a persistent artif

RippleKV: Cross-Layer KV Cache Allocation via Perturbation Propagation

Model ReleasesDGX agent

arXiv:2608.08684v1 Announce Type: cross Abstract: Long-context LLM inference is bottlenecked by KV cache memory, yet distributing a limited cache budget across layers remains challenging. Existing met

RISE-RL: Rubric-Informed Selective Exploration for Open-Ended Reinforcement Learning

Model ReleasesDGX agent

arXiv:2608.09123v1 Announce Type: new Abstract: Aligning Large Language Models (LLMs) for open-ended tasks is challenging because responses must satisfy multidimensional criteria without following a s

RotaryQuant: Fitting 120B MoE Models on Consumer Hardware via Fused Compressed-Space Attention

Model ReleasesDGX agent

arXiv:2608.08081v1 Announce Type: cross Abstract: Large mixture-of-experts (MoE) language models with 26--120 billion parameters exceed the memory capacity of consumer devices through three simultaneo

RouteGuard: Certifying Routing Gain in LLM Multi-Agent Systems When Complementarity Is Not Enough

Model ReleasesDGX agent

arXiv:2608.07583v1 Announce Type: cross Abstract: Multi-agent LLM systems route among model-backed advisors, yet a deployer rarely knows before shipping whether routing will help at all. Prevailing ro

Router Sensitivity Under Lightweight Fine-Tuning Identifies Prunable Experts in Mixture-of-Experts Models

Model ReleasesDGX agent

arXiv:2608.07890v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models decouple total parameters from per-token compute, but deployment still requires storing every expert. Recent theory sh

SafeSceneReason: A Multimodal Reasoning Benchmark Connecting Industrial Hazards with Accident Knowledge

Model ReleasesDGX agent

arXiv:2608.09230v1 Announce Type: new Abstract: Industrial-safety understanding requires more than detecting workers, equipment, and personal protective equipment. Models must also assess compliance,

SAGE: SLO-Aware Adaptive Retrieval for Production RAG Systems

Model ReleasesDGX agent

arXiv:2608.08237v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems in production operate under strict service level objectives (SLOs) on tail latency and infrastructure cos

SAIN: Structure-Aware Interactive Navigation with Active Dialogue Grounding for Mobile Robot

Model ReleasesDGX agent

arXiv:2608.09196v1 Announce Type: new Abstract: Most existing vision-language navigation tasks assume that instructions are complete and unambiguous. However, real-world robots often encounter natural

Same Question, Different Answer? Measuring and Mitigating Prompt Privilege for Equitable AI Access

Model ReleasesDGX agent

arXiv:2608.08942v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into healthcare, education, public services, and everyday decision making. They should provide

Sci-VBench: Evaluating Knowledge- and Reasoning-Intensive Video Generation in Science Domains

Model ReleasesDGX agent

arXiv:2608.09873v1 Announce Type: cross Abstract: We introduce Sci-VBench, a comprehensive benchmark for evaluating knowledge- and reasoning-intensive video generation across scientific domains. It co

SciTaRC: A Plan-Annotated Scientific Tabular QA Benchmark for Language Reasoning and Complex Computation

Model ReleasesDGX agent

arXiv:2603.08910v2 Announce Type: replace Abstract: We introduce SciTaRC, an expert-authored benchmark for question answering over scientific tables that targets composite, multi-step reasoning. To en

SCTD 3.0: Sonar Common Target Detection in the Wild - A Large-Scale, Multi-Scene Dataset from Real Marine Surveys

Model ReleasesDGX agent

arXiv:2608.08106v1 Announce Type: new Abstract: Synthetic Aperture Sonar (SAS) is core for wide-area detection of small underwater targets. However, large-scale, high-quality SAS datasets are scarce,

← Previous
1…89101112…369
Next →