AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Addressing Terminal Constraints in Data-Driven Demand Response Scheduling

DGX agent

arXiv:2605.14741v1 Announce Type: cross Abstract: Electrified chemical processes are incentivized by exposure to time-varying electricity markets to operate flexibly, but participating in demand respo

model-releasesarxiv-cs-ai
15 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Agentic Design of Compositional Descriptors via Autoresearch for Materials Science Applications

DGX agent

arXiv:2605.14671v1 Announce Type: cross Abstract: Autoresearch offers a flexible paradigm for automating scientific tasks, in which an AI agent proposes, implements, evaluates, and refines candidate s

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Agentic Recommender System with Hierarchical Belief-State Memory

DGX agent

arXiv:2605.14401v1 Announce Type: cross Abstract: Memory-augmented LLM agents have advanced personalized recommendation, yet existing approaches universally adopt flat memory representations that conf

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Agentic Systems as Boosting Weak Reasoning Models

DGX agent

arXiv:2605.14163v1 Announce Type: new Abstract: Can a committee of weak reasoning-model calls reach the performance of much stronger models? We study verifier-backed committee search as inference-time

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models

DGX agent

arXiv:2509.26100v2 Announce Type: replace Abstract: The rapid integration of Large Language Models (LLMs) into high-stakes domains necessitates reliable safety and compliance evaluation. However, exis

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

AgentTrap: Measuring Runtime Trust Failures in Third-Party Agent Skills

DGX agent

arXiv:2605.13940v1 Announce Type: cross Abstract: Third-party skills are becoming the package ecosystem for LLM agents. They package natural-language instructions, helper scripts, templates, documents

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

AI-assisted cultural heritage dissemination: Comparing NMT and glossary-augmented LLM translation in rock art documents

DGX agent

arXiv:2605.14679v1 Announce Type: cross Abstract: Cultural heritage institutions increasingly disseminate research and interpretive materials globally, but multilingual dissemination is constrained by

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

AnchorRoute: Human Motion Synthesis with Interval-Routed Sparse Contro

DGX agent

arXiv:2605.14716v1 Announce Type: cross Abstract: Sparse anchors provide a compact interface for human motion authoring: users specify a few root positions, planar trajectory samples, or body-point ta

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Are Agents Ready to Teach? A Multi-Stage Benchmark for Real-World Teaching Workflows

DGX agent

arXiv:2605.14322v1 Announce Type: new Abstract: Language agents are increasingly deployed in complex professional workflows, with tutoring emerging as a particularly high-stakes capability that remain

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

ARES-LSHADE: Autoresearch-Enhanced LSHADE with Memetic Polish for the GNBG Benchmark

DGX agent

arXiv:2605.13877v1 Announce Type: cross Abstract: We present ARES-LSHADE, a memetic differential-evolution variant submitted to the GECCO 2026 competition on LLM-designed evolutionary algorithms for t

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

ArGEnT: Arbitrary Geometry-encoded Transformer for Operator Learning

DGX agent

arXiv:2602.11626v2 Announce Type: replace-cross Abstract: Learning solution operators for systems with complex, varying geometries and parametric physical settings is a central challenge in scientific

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Asymmetric Generative Recommendation via Multi-Expert Projection and Multi-Faceted Hierarchical Quantization

DGX agent

arXiv:2605.14512v1 Announce Type: cross Abstract: Generative Recommendation (GenRec) models reformulate recommendation as a sequence generation task, representing items as discrete Semantic IDs used s

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Attention-Based Multimodal Survival Prediction with Cross-Modal Bilinear Fusion

DGX agent

arXiv:2605.13897v1 Announce Type: cross Abstract: We propose a novel multimodal deep learning framework for patient-level survival prediction, which integrates whole-slide histology features, RNA-seq

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

AttnGen: Attention-Guided Saliency Learning for Interpretable Genomic Sequence Classification

DGX agent

arXiv:2605.14073v1 Announce Type: cross Abstract: Deep neural networks have achieved strong performance in genomic sequence classification; however, relating their predictions to biologically meaningf

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Auditing Agent Harness Safety

DGX agent

arXiv:2605.14271v1 Announce Type: new Abstract: LLM agents increasingly run inside execution harnesses that dispatch tools, allocate resources, and route messages between specialized components. Howev

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Automated Construction of a Knowledge Graph of Nuclear Fusion Energy for Effective Elicitation and Retrieval of Information

DGX agent

arXiv:2504.07738v3 Announce Type: replace Abstract: In this document, we discuss a multi-step approach to automated construction of a knowledge graph, for structuring and representing domain-specific

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Beyond AI as Assistants: Toward Autonomous Discovery in Cosmology

DGX agent

arXiv:2605.14791v1 Announce Type: cross Abstract: Recent advances in artificial intelligence (AI) agents are pushing AI beyond tools toward autonomous scientific discovery. We discuss two complementar

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Beyond Binary: Reframing GUI Critique as Continuous Semantic Alignment

DGX agent

arXiv:2605.14311v1 Announce Type: cross Abstract: Test-Time Scaling (TTS), which samples multiple candidate actions and ranks them via a Critic Model, has emerged as a promising paradigm for generalis

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Beyond Mode-Seeking RL: Trajectory-Balance Post-Training for Diffusion Language Models

DGX agent

arXiv:2605.13935v1 Announce Type: cross Abstract: Diffusion language models are a promising alternative to autoregressive models, yet post-training methods for them largely adapt reward-maximizing obj

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

BiFedKD: Bidirectional Federated Knowledge Distillation Framework for Non-IID and Long-Tailed ECG Monitoring

DGX agent

arXiv:2605.14886v1 Announce Type: new Abstract: Electrocardiogram (ECG) monitoring in Internet of Medical Things (IoMT) networks is constrained by strict data-sharing regulations and privacy concerns.

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

BioHuman: Learning Biomechanical Human Representations from Video

DGX agent

arXiv:2605.14772v1 Announce Type: new Abstract: Understanding human motion beyond surface kinematics is crucial for motion analysis, rehabilitation, and injury risk assessment. However, progress in th

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

BiTrajDiff: Bidirectional Trajectory Generation with Diffusion Models for Offline Reinforcement Learning

DGX agent

arXiv:2506.05762v5 Announce Type: replace Abstract: Recent advances in offline Reinforcement Learning (RL) have proven that effective policy learning can benefit from imposing conservative constraints

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Breaking Dual Bottlenecks: Evolving Unified Multimodal Models into Self-Adaptive Interleaved Visual Reasoners

DGX agent

arXiv:2605.14709v1 Announce Type: new Abstract: Recent unified models integrate multimodal understanding and generation within a single framework. However, an 'understanding-generation gap' persists,

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Can Visual Mamba Improve AI-Generated Image Detection? An In-Depth Investigation

DGX agent

arXiv:2605.14799v1 Announce Type: new Abstract: In recent years, computer vision has witnessed remarkable progress, fueled by the development of innovative architectures such as Convolutional Neural N

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Cattle Trade: A Multi-Agent Benchmark for LLM Bluffing, Bidding, and Bargaining

DGX agent

arXiv:2605.14537v1 Announce Type: new Abstract: We introduce extsc{Cattle Trade, a multi-agent benchmark for evaluating large language models (LLMs) as agents in strategic reasoning under imperfect in

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

CausalReasoningBenchmark: A Real-World Benchmark for Disentangled Evaluation of Causal Identification and Estimation

DGX agent

arXiv:2602.20571v2 Announce Type: replace Abstract: Many benchmarks for automated causal inference evaluate a system's performance based on a single numerical output, such as an Average Treatment Effe

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Chain-of-Procedure: Hierarchical Visual-Language Reasoning for Procedural QA

DGX agent

arXiv:2605.14928v1 Announce Type: new Abstract: Recent advances in vision-language models (VLMs) have achieved impressive results on standard image-text tasks, yet their potential for visual procedure

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Chinese Short-Form Creative Content Generation via Explanation-Oriented Multi-Objective Optimization

DGX agent

arXiv:2511.15408v2 Announce Type: replace-cross Abstract: Chinese demonstrates high semantic compactness and rich metaphorical expressiveness, enabling limited text to convey dense meanings while incr

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

CineMesh4D: Personalized 4D Whole Heart Reconstruction from Sparse Cine MRI

DGX agent

arXiv:2605.13994v1 Announce Type: cross Abstract: Accurate 3D+t whole-heart mesh reconstruction from cine MRI is a clinically crucial yet technically challenging task. The difficulty of this task aris

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents

DGX agent

arXiv:2605.14133v1 Announce Type: new Abstract: Interactive agent benchmarks face a tension between scalable construction and realistic workflow evaluation. Hand-authored tasks are expensive to extend

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

CLOVER: Closed-Loop Value Estimation & Ranking for End-to-End Autonomous Driving Planning

DGX agent

arXiv:2605.15120v1 Announce Type: cross Abstract: End-to-end autonomous driving planners are commonly trained by imitating a single logged trajectory, yet evaluated by rule-based planning metrics that

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

CoCoEdit: Content-Consistent Image Editing via Region Regularized Reinforcement Learning

DGX agent

arXiv:2602.14068v2 Announce Type: replace Abstract: Image editing has achieved impressive results with the development of large-scale generative models. However, existing models mainly focus on the ed

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Cognitive-Uncertainty Guided Knowledge Distillation for Accurate Classification of Student Misconceptions

DGX agent

arXiv:2605.14752v1 Announce Type: cross Abstract: Accurately identifying student misconceptions is crucial for personalized education but faces three challenges: (1) data scarcity with long-tail distr

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Collider-Bench: Benchmarking AI Agents with Particle Physics Analysis Reproduction

DGX agent

arXiv:2605.13950v1 Announce Type: cross Abstract: Autonomous language-model agents are increasingly evaluated on long-horizon tool-use tasks, but existing benchmarks rarely capture the complexity and

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Communication-Efficient Federated Fine-Tuning

DGX agent

arXiv:2505.04535v3 Announce Type: replace Abstract: Federated Learning (FL) enables the utilization of vast, previously inaccessible data sources. At the same time, pre-trained Language Models (LMs) h

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Correctness-Aware Repository Filtering Under Maximum Effective Context Window Constraints

DGX agent

arXiv:2605.14362v1 Announce Type: cross Abstract: Context window efficiency is a practical constraint in large language model (LLM)-based developer tools. Paulsen [12] shows that all tested models deg

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

CounselBench: A Large-Scale Expert Evaluation and Adversarial Benchmarking of Large Language Models in Mental Health Question Answering

DGX agent

arXiv:2506.08584v4 Announce Type: replace Abstract: Medical question answering (QA) benchmarks often focus on multiple-choice or fact-based tasks, leaving open-ended answers to real patient questions

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

CRANE: Constrained Reasoning Injection for Code Agents via Nullspace Editing

DGX agent

arXiv:2605.14084v1 Announce Type: cross Abstract: Code agents must both reason over long-horizon repository state and obey strict tool-use protocols. In paired Instruct/Thinking checkpoints, these cap

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Critic-Driven Voronoi-Quantization for Distilling Deep RL Policies to Explainable Models

DGX agent

arXiv:2605.14897v1 Announce Type: cross Abstract: Despite many successful attempts at explaining Deep Reinforcement Learning policies using distillation, it remains difficult to balance the performanc

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

CUICurate: A GraphRAG-based Framework for Automated Clinical Concept Curation for NLP applications

DGX agent

arXiv:2602.17949v2 Announce Type: replace-cross Abstract: Background: Clinical named entity recognition tools commonly map free text to Unified Medical Language System (UMLS) Concept Unique Identifier

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves

DGX agent

arXiv:2605.14068v1 Announce Type: new Abstract: We introduce CurveBench, a benchmark for hierarchical topological reasoning from visual input. CurveBench consists of extbf{756 images} of pairwise non-

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Darwin Family: MRI-Trust-Weighted Evolutionary Merging for Training-Free Scaling of Language-Model Reasoning

DGX agent

arXiv:2605.14386v1 Announce Type: cross Abstract: We present Darwin Family, a framework for training-free evolutionary merging of large language models via gradient-free weight-space recombination. We

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Data-Augmented Game Starts for Accelerating Self-Play Exploration in Imperfect Information Games

DGX agent

arXiv:2605.14379v1 Announce Type: cross Abstract: Finding approximate equilibria for large-scale imperfect-information competitive games such as StarCraft, Dota, and CounterStrike remains computationa

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Deceive, Detect, and Disclose: Large Language Models Play Mini-Mafia

DGX agent

arXiv:2509.23023v3 Announce Type: replace Abstract: Large language models are increasingly deployed in multi-agent settings whose outcomes hinge on social intelligence, motivating evaluations of their

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Deep Image Segmentation via Discriminant Feature Learning

DGX agent

arXiv:2605.14609v1 Announce Type: new Abstract: Accurate image segmentation remains challenging, particularly in generating sharp, confident boundaries. While modern architectures have advanced the fi

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Denoising-GS: Gaussian Splatting with Spatial-aware Denoising

DGX agent

arXiv:2605.14880v1 Announce Type: new Abstract: Recent advances in 3D Gaussian Splatting (3DGS) have achieved remarkable success in high-fidelity Novel View Synthesis (NVS), yet the optimization proce

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Derivation Prompting: A Logic-Based Method for Improving Retrieval-Augmented Generation

DGX agent

arXiv:2605.14053v1 Announce Type: cross Abstract: The application of Large Language Models to Question Answering has shown great promise, but important challenges such as hallucinations and erroneous

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Descriptor: Distance-Annotated Traffic Perception Question Answering (DTPQA)

DGX agent

arXiv:2511.13397v2 Announce Type: replace-cross Abstract: The remarkable progress of Vision-Language Models (VLMs) on a variety of tasks has raised interest in their application to automated driving.

model-releasesarxiv-cs-ai
15 May 2026
← Previous
1…236237238239240…361
Next →