AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,585 results
15 May 2026

3D Skew-Normal Splatting

Model ReleasesDGX agent

arXiv:2605.15010v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has emerged as a leading representation for real-time novel view synthesis and been widely adopted in various downstream ap

40 Grok Build agents tearing through C code in parallel. All supervised by DAD. DAD is a lightweight autonomous tmux supervisor for long-run…

Model ReleasesDGX agent

40 Grok Build agents tearing through C code in parallel. All supervised by DAD. DAD is a lightweight autonomous tmux supervisor for long-running Grok tasks. Built entirely in Grok Build. /dad 'your ob

A Benchmark for Early-stage Parkinson's Disease Detection from Speech

Model ReleasesDGX agent

arXiv:2605.14066v1 Announce Type: cross Abstract: Early-stage Parkinson's disease (EarlyPD) detection from speech is clinically meaningful yet underexplored, and published results are hard to compare


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

A Calculus-Based Framework for Determining Vocabulary Size in End-to-End ASR

Model ReleasesDGX agent

arXiv:2605.14427v1 Announce Type: new Abstract: In hybrid automatic speech recognition (ASR) systems, the vocabulary size is unambiguous, typically determined by the number of phones, bi-phones, or tr

A Deterministic Agentic Workflow for HS Tariff Classification: Multi-Dimensional Rule Reasoning with Interpretable Decisions

Model ReleasesDGX agent

arXiv:2605.14857v1 Announce Type: new Abstract: Harmonized System (HS) tariff classification is a high-stakes, expert-level task in which a free-form product description must be mapped to a specific s

A Large Language Model Based Pipeline for Review of Systems Entity Recognition from Clinical Notes

Model ReleasesDGX agent

arXiv:2506.11067v3 Announce Type: replace Abstract: Objective: Develop a cost-effective, large language model (LLM)-based pipeline for automatically extracting Review of Systems (ROS) entities from cl

A milestone for Pearl Research Labs: our first major enterprise partnership is live with Together AI. @togethercompute’s inference platform …

Model ReleasesDGX agent

A milestone for Pearl Research Labs: our first major enterprise partnership is live with Together AI. @togethercompute’s inference platform is an ideal demonstration of @prlnet's value proposition — O

A Minimal Agent for Automated Theorem Proving

Model ReleasesDGX agent

arXiv:2602.24273v3 Announce Type: replace Abstract: We propose a minimal agentic baseline that enables systematic comparison across different AI-based theorem prover architectures. This design impleme

A new set of open-weight models is topping the leaderboard for document understanding 🔥 INF just released two models: Infinity-Parser2-Pro …

Model ReleasesDGX agent

A new set of open-weight models is topping the leaderboard for document understanding 🔥 INF just released two models: Infinity-Parser2-Pro (35B) and Infinity-Parser2-Flash (2B) that top our @huggingfa

A Non-Destructive Methodological Framework for Modernizing Legacy Clinical Reporting Systems for AI-Driven Pharmacoinformatics: A SAS Case Study

Model ReleasesDGX agent

arXiv:2605.13905v1 Announce Type: cross Abstract: Drug development and pharmacovigilance are frequently bottlenecked by legacy clinical reporting pipelines. These monolithic systems encode regulatory-

A Picture is Worth a Thousand Words? An Empirical Study of Aggregation Strategies for Visual Financial Document Retrieval

Model ReleasesDGX agent

arXiv:2605.14581v1 Announce Type: cross Abstract: Visual RAG has offered an alternative to traditional RAG. It treats documents as images and uses vision encoders to obtain vision patch tokens. Howeve

A Problem-Oriented Taxonomy of Evaluation Metrics for Time Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2511.18739v2 Announce Type: replace Abstract: Time series anomaly detection is widely used in IoT and cyber-physical systems, yet its evaluation remains challenging due to diverse application ob

A Survey on Data-Dependent Worst-Case Generalization Bounds

Model ReleasesDGX agent

arXiv:2605.13913v1 Announce Type: cross Abstract: Deep neural networks generalize well despite being heavily overparameterized, in apparent contradiction with classical learning theory based on unifor

A Tutorial on Cognitive Biases in Agentic AI-Driven 6G Autonomous Networks

Model ReleasesDGX agent

arXiv:2510.19973v4 Announce Type: replace-cross Abstract: The path to higher network autonomy in 6G lies beyond the mere optimization of key performance indicators (KPIs), requiring systems that perce

ACE-LoRA: Adaptive Orthogonal Decoupling for Continual Image Editing

Model ReleasesDGX agent

arXiv:2605.14948v1 Announce Type: new Abstract: State-of-the-art diffusion models often rely on parameter-efficient fine-tuning to perform specialized image editing tasks. However, real-world applicat

Action-Inspired Generative Models

Model ReleasesDGX agent

arXiv:2605.14631v1 Announce Type: cross Abstract: We introduce Action-Inspired Generative Models (AGMs), a dual-network generative framework motivated by the observation that existing bridge-matching

Addressing Terminal Constraints in Data-Driven Demand Response Scheduling

Model ReleasesDGX agent

arXiv:2605.14741v1 Announce Type: cross Abstract: Electrified chemical processes are incentivized by exposure to time-varying electricity markets to operate flexibly, but participating in demand respo

Agentic Design of Compositional Descriptors via Autoresearch for Materials Science Applications

Model ReleasesDGX agent

arXiv:2605.14671v1 Announce Type: cross Abstract: Autoresearch offers a flexible paradigm for automating scientific tasks, in which an AI agent proposes, implements, evaluates, and refines candidate s

Agentic Recommender System with Hierarchical Belief-State Memory

Model ReleasesDGX agent

arXiv:2605.14401v1 Announce Type: cross Abstract: Memory-augmented LLM agents have advanced personalized recommendation, yet existing approaches universally adopt flat memory representations that conf

Agentic Systems as Boosting Weak Reasoning Models

Model ReleasesDGX agent

arXiv:2605.14163v1 Announce Type: new Abstract: Can a committee of weak reasoning-model calls reach the performance of much stronger models? We study verifier-backed committee search as inference-time

AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models

Model ReleasesDGX agent

arXiv:2509.26100v2 Announce Type: replace Abstract: The rapid integration of Large Language Models (LLMs) into high-stakes domains necessitates reliable safety and compliance evaluation. However, exis

AgentTrap: Measuring Runtime Trust Failures in Third-Party Agent Skills

Model ReleasesDGX agent

arXiv:2605.13940v1 Announce Type: cross Abstract: Third-party skills are becoming the package ecosystem for LLM agents. They package natural-language instructions, helper scripts, templates, documents

AI-assisted cultural heritage dissemination: Comparing NMT and glossary-augmented LLM translation in rock art documents

Model ReleasesDGX agent

arXiv:2605.14679v1 Announce Type: cross Abstract: Cultural heritage institutions increasingly disseminate research and interpretive materials globally, but multilingual dissemination is constrained by

AI radio hosts demonstrate why AI can’t be trusted alone

Model ReleasesDGX agent

Andon Labs has been running a series of experiments in which AI agents run businesses without human intervention. Its latest is a quartet of radio stations run by some of the most popular AI models ou

AnchorRoute: Human Motion Synthesis with Interval-Routed Sparse Contro

Model ReleasesDGX agent

arXiv:2605.14716v1 Announce Type: cross Abstract: Sparse anchors provide a compact interface for human motion authoring: users specify a few root positions, planar trajectory samples, or body-point ta

Anthropic just went after the 44% of U.S. GDP that enterprise AI has mostly ignored. Claude for Small Business launched this week with 15 pr…

Model ReleasesDGX agent

Anthropic just went after the 44% of U.S. GDP that enterprise AI has mostly ignored. Claude for Small Business launched this week with 15 prebuilt agentic workflows and 15 skills connected directly in

Are Agents Ready to Teach? A Multi-Stage Benchmark for Real-World Teaching Workflows

Model ReleasesDGX agent

arXiv:2605.14322v1 Announce Type: new Abstract: Language agents are increasingly deployed in complex professional workflows, with tutoring emerging as a particularly high-stakes capability that remain

ARES-LSHADE: Autoresearch-Enhanced LSHADE with Memetic Polish for the GNBG Benchmark

Model ReleasesDGX agent

arXiv:2605.13877v1 Announce Type: cross Abstract: We present ARES-LSHADE, a memetic differential-evolution variant submitted to the GECCO 2026 competition on LLM-designed evolutionary algorithms for t

ArGEnT: Arbitrary Geometry-encoded Transformer for Operator Learning

Model ReleasesDGX agent

arXiv:2602.11626v2 Announce Type: replace-cross Abstract: Learning solution operators for systems with complex, varying geometries and parametric physical settings is a central challenge in scientific

Asymmetric Generative Recommendation via Multi-Expert Projection and Multi-Faceted Hierarchical Quantization

Model ReleasesDGX agent

arXiv:2605.14512v1 Announce Type: cross Abstract: Generative Recommendation (GenRec) models reformulate recommendation as a sequence generation task, representing items as discrete Semantic IDs used s

Attention-Based Multimodal Survival Prediction with Cross-Modal Bilinear Fusion

Model ReleasesDGX agent

arXiv:2605.13897v1 Announce Type: cross Abstract: We propose a novel multimodal deep learning framework for patient-level survival prediction, which integrates whole-slide histology features, RNA-seq

AttnGen: Attention-Guided Saliency Learning for Interpretable Genomic Sequence Classification

Model ReleasesDGX agent

arXiv:2605.14073v1 Announce Type: cross Abstract: Deep neural networks have achieved strong performance in genomic sequence classification; however, relating their predictions to biologically meaningf

Auditing Agent Harness Safety

Model ReleasesDGX agent

arXiv:2605.14271v1 Announce Type: new Abstract: LLM agents increasingly run inside execution harnesses that dispatch tools, allocate resources, and route messages between specialized components. Howev

Automated Construction of a Knowledge Graph of Nuclear Fusion Energy for Effective Elicitation and Retrieval of Information

Model ReleasesDGX agent

arXiv:2504.07738v3 Announce Type: replace Abstract: In this document, we discuss a multi-step approach to automated construction of a knowledge graph, for structuring and representing domain-specific

Beyond AI as Assistants: Toward Autonomous Discovery in Cosmology

Model ReleasesDGX agent

arXiv:2605.14791v1 Announce Type: cross Abstract: Recent advances in artificial intelligence (AI) agents are pushing AI beyond tools toward autonomous scientific discovery. We discuss two complementar

Beyond Binary: Reframing GUI Critique as Continuous Semantic Alignment

Model ReleasesDGX agent

arXiv:2605.14311v1 Announce Type: cross Abstract: Test-Time Scaling (TTS), which samples multiple candidate actions and ranks them via a Critic Model, has emerged as a promising paradigm for generalis

Beyond Mode-Seeking RL: Trajectory-Balance Post-Training for Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.13935v1 Announce Type: cross Abstract: Diffusion language models are a promising alternative to autoregressive models, yet post-training methods for them largely adapt reward-maximizing obj

BiFedKD: Bidirectional Federated Knowledge Distillation Framework for Non-IID and Long-Tailed ECG Monitoring

Model ReleasesDGX agent

arXiv:2605.14886v1 Announce Type: new Abstract: Electrocardiogram (ECG) monitoring in Internet of Medical Things (IoMT) networks is constrained by strict data-sharing regulations and privacy concerns.

BioHuman: Learning Biomechanical Human Representations from Video

Model ReleasesDGX agent

arXiv:2605.14772v1 Announce Type: new Abstract: Understanding human motion beyond surface kinematics is crucial for motion analysis, rehabilitation, and injury risk assessment. However, progress in th

BiTrajDiff: Bidirectional Trajectory Generation with Diffusion Models for Offline Reinforcement Learning

Model ReleasesDGX agent

arXiv:2506.05762v5 Announce Type: replace Abstract: Recent advances in offline Reinforcement Learning (RL) have proven that effective policy learning can benefit from imposing conservative constraints

Breaking Dual Bottlenecks: Evolving Unified Multimodal Models into Self-Adaptive Interleaved Visual Reasoners

Model ReleasesDGX agent

arXiv:2605.14709v1 Announce Type: new Abstract: Recent unified models integrate multimodal understanding and generation within a single framework. However, an 'understanding-generation gap' persists,

Can Visual Mamba Improve AI-Generated Image Detection? An In-Depth Investigation

Model ReleasesDGX agent

arXiv:2605.14799v1 Announce Type: new Abstract: In recent years, computer vision has witnessed remarkable progress, fueled by the development of innovative architectures such as Convolutional Neural N

Cattle Trade: A Multi-Agent Benchmark for LLM Bluffing, Bidding, and Bargaining

Model ReleasesDGX agent

arXiv:2605.14537v1 Announce Type: new Abstract: We introduce extsc{Cattle Trade, a multi-agent benchmark for evaluating large language models (LLMs) as agents in strategic reasoning under imperfect in

CausalReasoningBenchmark: A Real-World Benchmark for Disentangled Evaluation of Causal Identification and Estimation

Model ReleasesDGX agent

arXiv:2602.20571v2 Announce Type: replace Abstract: Many benchmarks for automated causal inference evaluate a system's performance based on a single numerical output, such as an Average Treatment Effe

Chain-of-Procedure: Hierarchical Visual-Language Reasoning for Procedural QA

Model ReleasesDGX agent

arXiv:2605.14928v1 Announce Type: new Abstract: Recent advances in vision-language models (VLMs) have achieved impressive results on standard image-text tasks, yet their potential for visual procedure

Chinese Short-Form Creative Content Generation via Explanation-Oriented Multi-Objective Optimization

Model ReleasesDGX agent

arXiv:2511.15408v2 Announce Type: replace-cross Abstract: Chinese demonstrates high semantic compactness and rich metaphorical expressiveness, enabling limited text to convey dense meanings while incr

CineMesh4D: Personalized 4D Whole Heart Reconstruction from Sparse Cine MRI

Model ReleasesDGX agent

arXiv:2605.13994v1 Announce Type: cross Abstract: Accurate 3D+t whole-heart mesh reconstruction from cine MRI is a clinically crucial yet technically challenging task. The difficulty of this task aris

Claude Code's product lead talks usage limits, transparency, and the 'lean harness'

Model ReleasesDGX agent

Claude Code's product lead addresses how Anthropic tunes the 'harness' (the structural layer around the model) for each new model release to optimize performance and reduce verbosity. The company comm

ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents

Model ReleasesDGX agent

arXiv:2605.14133v1 Announce Type: new Abstract: Interactive agent benchmarks face a tension between scalable construction and realistic workflow evaluation. Hand-authored tasks are expensive to extend

CLOVER: Closed-Loop Value Estimation & Ranking for End-to-End Autonomous Driving Planning

Model ReleasesDGX agent

arXiv:2605.15120v1 Announce Type: cross Abstract: End-to-end autonomous driving planners are commonly trained by imitating a single logged trajectory, yet evaluated by rule-based planning metrics that

CoCoEdit: Content-Consistent Image Editing via Region Regularized Reinforcement Learning

Model ReleasesDGX agent

arXiv:2602.14068v2 Announce Type: replace Abstract: Image editing has achieved impressive results with the development of large-scale generative models. However, existing models mainly focus on the ed

Cognitive-Uncertainty Guided Knowledge Distillation for Accurate Classification of Student Misconceptions

Model ReleasesDGX agent

arXiv:2605.14752v1 Announce Type: cross Abstract: Accurately identifying student misconceptions is crucial for personalized education but faces three challenges: (1) data scarcity with long-tail distr

Collider-Bench: Benchmarking AI Agents with Particle Physics Analysis Reproduction

Model ReleasesDGX agent

arXiv:2605.13950v1 Announce Type: cross Abstract: Autonomous language-model agents are increasingly evaluated on long-horizon tool-use tasks, but existing benchmarks rarely capture the complexity and

Communication-Efficient Federated Fine-Tuning

Model ReleasesDGX agent

arXiv:2505.04535v3 Announce Type: replace Abstract: Federated Learning (FL) enables the utilization of vast, previously inaccessible data sources. At the same time, pre-trained Language Models (LMs) h

Correctness-Aware Repository Filtering Under Maximum Effective Context Window Constraints

Model ReleasesDGX agent

arXiv:2605.14362v1 Announce Type: cross Abstract: Context window efficiency is a practical constraint in large language model (LLM)-based developer tools. Paulsen [12] shows that all tested models deg

CounselBench: A Large-Scale Expert Evaluation and Adversarial Benchmarking of Large Language Models in Mental Health Question Answering

Model ReleasesDGX agent

arXiv:2506.08584v4 Announce Type: replace Abstract: Medical question answering (QA) benchmarks often focus on multiple-choice or fact-based tasks, leaving open-ended answers to real patient questions

CRANE: Constrained Reasoning Injection for Code Agents via Nullspace Editing

Model ReleasesDGX agent

arXiv:2605.14084v1 Announce Type: cross Abstract: Code agents must both reason over long-horizon repository state and obey strict tool-use protocols. In paired Instruct/Thinking checkpoints, these cap

Critic-Driven Voronoi-Quantization for Distilling Deep RL Policies to Explainable Models

Model ReleasesDGX agent

arXiv:2605.14897v1 Announce Type: cross Abstract: Despite many successful attempts at explaining Deep Reinforcement Learning policies using distillation, it remains difficult to balance the performanc

CUICurate: A GraphRAG-based Framework for Automated Clinical Concept Curation for NLP applications

Model ReleasesDGX agent

arXiv:2602.17949v2 Announce Type: replace-cross Abstract: Background: Clinical named entity recognition tools commonly map free text to Unified Medical Language System (UMLS) Concept Unique Identifier

CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves

Model ReleasesDGX agent

arXiv:2605.14068v1 Announce Type: new Abstract: We introduce CurveBench, a benchmark for hierarchical topological reasoning from visual input. CurveBench consists of extbf{756 images} of pairwise non-

← Previous
1…246247248249250…377
Next →