AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,797 results
Model Releases

Scaling GraphLLM with Bilevel-Optimized Sparse Querying

DGX agent

arXiv:2602.09038v2 Announce Type: replace-cross Abstract: LLMs have recently shown strong potential in enhancing node-level tasks on text-attributed graphs (TAGs) by providing explanation features. Ho

model-releasesarxiv-cs-ai
27 May 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Scheduled Style Injection: Expanding the Style-Content Pareto Frontier in Training-Free Diffusion-based Style Transfer

DGX agent

arXiv:2605.26538v1 Announce Type: new Abstract: Style transfer with pre-trained diffusion models has advanced rapidly, but a core question remains underexplored: where in the model should style inject

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence

DGX agent

arXiv:2605.26340v1 Announce Type: new Abstract: Autonomous research agents produce competitive solutions and professional-looking manuscripts, yet their outputs contain verifiability failures undetect

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

SEAL: Self-Evolving Agentic Learning for Conversational Question Answering over Knowledge Graphs

DGX agent

arXiv:2512.04868v2 Announce Type: replace-cross Abstract: Knowledge-based conversational question answering (KBCQA) confronts persistent challenges in resolving coreference, modeling contextual depend

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks?

DGX agent

arXiv:2605.26548v1 Announce Type: cross Abstract: Large language models (LLMs) now support automated software security tasks, including vulnerability discovery and proof-of-concept (PoC) generation. E

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

SeDT: Sentence-Transformer Decision-Transformer Conditioning for Multi-Turn Conversation Reliability

DGX agent

arXiv:2605.26788v1 Announce Type: cross Abstract: Large language models (LLMs) achieve impressive performance when a task is fully specified in a single turn, yet the same models lose up to 39% of tha

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Seeing vs. Believing: Evaluating the Language Bias of Open-Source MLLMs in Counter-Intuitive Scenes

DGX agent

arXiv:2601.07737v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable performance in mainstream visual understanding tasks, but their ability

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Self-Ensembling Vision-Language Models for Chart Data Extraction

DGX agent

arXiv:2605.27298v1 Announce Type: new Abstract: Charts effectively convey quantitative information, but the underlying data are often locked in image form, hindering reuse and analysis. Manually digit

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Self-Verified Distillation: Your Language Model Is Secretly Its Own Synthetic Data Pipeline

DGX agent

arXiv:2605.26132v1 Announce Type: new Abstract: Can post-trained large language models (LLMs) further improve themselves using only unlabeled prompts, without external teachers or feedback from tools?

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Sentinel: Embodied Cooperative Spatial Reasoning and Planning

DGX agent

arXiv:2605.26239v1 Announce Type: new Abstract: In this work, we study Cooperative Spatial Intelligence, the ability of decentralized embodied agents to coordinate effectively under dynamic environmen

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Separating Semantic Competition from Context Length in RAG Reading

DGX agent

arXiv:2605.27294v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) systems can respond incorrectly even when the correct passage was retrieved. The model must still read the retrieve

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Shedding Light on Dark Matter at the LHC with Machine Learning

DGX agent

arXiv:2509.15121v2 Announce Type: replace-cross Abstract: We investigate a WIMP dark matter (DM) candidate in the form of a singlino-dominated lightest supersymmetric particle (LSP) within the Z_3-sym

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Shopping Companion: A Memory-Augmented LLM Agent for Real-World E-Commerce Tasks

DGX agent

arXiv:2603.14864v2 Announce Type: replace Abstract: In e-commerce, LLM agents show promise for shopping tasks such as recommendations, budget management, and bundle deals, where accurately capturing u

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

SilIF: Silhouette-Augmented Isolation Forest for Unsupervised Transaction Fraud Detection

DGX agent

arXiv:2605.26135v1 Announce Type: new Abstract: Unsupervised anomaly detection is widely used in transaction fraud detection where labels are scarce. Isolation Forest (IF) is among the most popular cl

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning

DGX agent

arXiv:2603.28730v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have shown impressive capabilities across diverse tasks, motivating efforts to leverage these models to supervis

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Someone just asked me “What I don’t get is why you see it as some kind of failure that AI labs are now using harnesses and neurosymbolic too…

DGX agent

Someone just asked me “What I don’t get is why you see it as some kind of failure that AI labs are now using harnesses and neurosymbolic tools…?” I don’t! As I argued in my newsletter on Claude Code (

model-releasesgary-marcus--x
27 May 2026
Model Releases

SONAR-LLM: Autoregressive Transformer that Thinks in Sentence Embeddings and Speaks in Tokens

DGX agent

arXiv:2508.05305v2 Announce Type: replace Abstract: The recently proposed Large Concept Model (LCM) generates text by predicting a sequence of sentence-level embeddings and training with either mean-s

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

SpaceVista: All-Scale Visual Spatial Reasoning from mm to km

DGX agent

arXiv:2510.09606v2 Announce Type: replace Abstract: With the current surge in spatial reasoning explorations, researchers have made significant progress in understanding indoor scenes, but still strug

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

SpatialBench: Is Your Spatial Foundation Model an All-Round Player?

DGX agent

arXiv:2605.27367v1 Announce Type: new Abstract: While spatial foundation models have demonstrated impressive performance on standard datasets, a critical question remains: are they truly all-round pla

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Stability Implies Redundancy: Delta Attention Selective Halting for Efficient Long-Context Prefilling

DGX agent

arXiv:2604.18103v2 Announce Type: replace Abstract: Prefilling computational costs pose a significant bottleneck for Large Language Models (LLMs) and Large Multimodal Models (LMMs) in long-context set

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

SteelDS: A High-Resolution Video Dataset of E40 Steel Scrap for Object Detection and Instance Segmentation

DGX agent

arXiv:2605.26682v1 Announce Type: cross Abstract: This dataset provides high-resolution, annotated video sequences of shredded E40-grade steel and copper scrap on a conveyor belt. Captured in a contro

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Stochastic global optimization of continuous functions via random walks on Grassmannians

DGX agent

arXiv:2605.14151v1 Announce Type: cross Abstract: We introduce a stochastic global optimization method based on random walks on Grassmannian manifolds. To minimize a continuous objective ell:R^drighta

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Strategies for Guiding LLMs to Use Software Design Patterns: A Case of Singleton

DGX agent

arXiv:2605.26898v1 Announce Type: cross Abstract: Large Language Models (LLMs) can generate functional source code from natural-language prompts, but often fail to consistently follow higher-level arc

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Stronger models do not always need lighter harnesses. Everyone believes more structured harnesses universally improve reliability, and that …

DGX agent

Stronger models do not always need lighter harnesses. Everyone believes more structured harnesses universally improve reliability, and that higher-capability models need proportionally less structural

model-releasesdair-ai--x
27 May 2026
Model Releases

Structured Relational Reasoning for Group Activity Assessment

DGX agent

arXiv:2508.07996v2 Announce Type: replace Abstract: Group Activity Detection (GAD) involves recognizing social groups and their collective behaviors in videos. Vision Foundation Models (VFMs), like DI

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

SWE-Adept: An LLM-Based Agentic Framework for Deep Codebase Analysis and Structured Issue Resolution

DGX agent

arXiv:2603.01327v2 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong performance on self-contained programming tasks. However, they still struggle with repository-leve

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

TADDLE: A Tool-Augmented Agent for Detecting Deficient LLM-Generated Peer Reviews

DGX agent

arXiv:2605.26911v1 Announce Type: new Abstract: LLM-generated peer reviews are increasingly common at major venues, yet their deficiencies are hard to detect because they are uniformly fluent and well

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Temporal Simultaneity Predicts Annotation Quality in Sentiment Corpora

DGX agent

arXiv:2605.27239v1 Announce Type: new Abstract: Annotation quality is difficult to sustain when campaigns span weeks or months with small annotator pools. We present a Setswana sentiment dataset of 3,

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

The Bridge-Garden Dilemma in LLM Distillation: Why Mixing Hard and Soft Labels Works

DGX agent

arXiv:2605.26246v1 Announce Type: new Abstract: Knowledge distillation (KD) transfers knowledge from a large teacher model to a smaller student. In language modeling, the student is trained either on

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

The Compressive Knowledge Graph Hypothesis: Which Graph Facts Matter for Scientific Hypothesis Generation?

DGX agent

arXiv:2605.27176v1 Announce Type: new Abstract: Knowledge graphs (KGs) can provide structured scientific context to language models, but it remains unclear which graph facts actually shape the generat

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

The MiniMax M2 series was one of the most widely used open-weight LLM series earlier this year. Now, we got a technical report with some int…

DGX agent

The MiniMax M2 series was one of the most widely used open-weight LLM series earlier this year. Now, we got a technical report with some interesting tidbits. I summarized some of them below: 1. Full a

model-releasessebastian-raschka--x
27 May 2026
Model Releases

The Stability of Singular Distribution: A Spectral Perspective on the Two-Phase Dynamics of Language Model Pre-training

DGX agent

arXiv:2605.26489v1 Announce Type: new Abstract: Large language model pre-training typically exhibits a two-phase trajectory: a fast initial loss drop followed by a prolonged slow improvement. We ident

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Timestep-Aware SVDQuant-GPTQ for W4A4 Quantization of Wan2.2-I2V

DGX agent

arXiv:2605.27003v1 Announce Type: cross Abstract: W4A4 quantization of large video diffusion Transformers offers substantial memory savings but is hindered by two main challenges: sparse large-magnitu

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Today we're announcing ESMFold2, an open scientific engine to power prediction, design, and discovery across protein biology. The new model …

DGX agent

Today we're announcing ESMFold2, an open scientific engine to power prediction, design, and discovery across protein biology. The new model delivers state of the art performance on protein interaction

model-releasesyann-lecun--x
27 May 2026
Model Releases

Tool-Schema Compression Enables Agentic RAG Under Constrained Context Budgets

DGX agent

arXiv:2605.26165v1 Announce Type: cross Abstract: Agentic RAG systems that equip language models with dozens to hundreds of tool definitions face a critical resource conflict: tool schemas consume the

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Towards Error-Free EHRs: Reasoning-Intensive Consistency Verification Between Clinical Notes and Structured Tables in Electronic Health Records

DGX agent

arXiv:2605.26463v1 Announce Type: cross Abstract: Data consistency between unstructured clinical notes and structured tables in Electronic Health Records (EHRs) is essential for patient safety and cli

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Towards Generalization-Oriented Models for Vehicle Routing Problems with Mixture-of-Experts

DGX agent

arXiv:2605.26776v1 Announce Type: cross Abstract: In recent years, Deep Reinforcement Learning (DRL) has achieved substantial progress on Vehicle Routing Problems (VRPs). However, existing DRL-based m

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

TowerMind: A Tower Defence Game Learning Environment and Benchmark for LLM as Agents

DGX agent

arXiv:2601.05899v2 Announce Type: replace Abstract: Recent breakthroughs in Large Language Models (LLMs) have positioned them as a promising paradigm for agents, with long-term planning and decision-m

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Traceable Knowledge Graph Reasoning Enables LLM-Assisted Decision Support for Industrial VOCs in the Steel Industry

DGX agent

arXiv:2605.27071v1 Announce Type: new Abstract: Key knowledge for steel-industry volatile organic compounds (VOCs) governance is scattered across unstructured scientific literature, making it difficul

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Trust Region Q Adjoint Matching

DGX agent

arXiv:2605.27079v1 Announce Type: cross Abstract: Off-policy reinforcement learning of pretrained flow policies remains challenging due to the instability of optimization arising from the multi-step s

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Two-Parameter Flows for Learning Population Dynamics of Physical Systems

DGX agent

arXiv:2605.26285v1 Announce Type: new Abstract: This work addresses the problem of learning the dynamics of high-dimensional probability densities over time using unlabeled samples, without assuming a

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Underwater360: Reconstructing Underwater Scenes from Panoramic Images with Omnidirectional Gaussian Splatting

DGX agent

arXiv:2605.26447v1 Announce Type: new Abstract: Underwater scene reconstruction is essential for immersive exploration of aquatic environments, yet remains challenging due to complex participating-med

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

UnityMAS-O: A General RL Optimization Framework for LLM-Based Multi-Agent Systems

DGX agent

arXiv:2605.26646v1 Announce Type: new Abstract: LLM-based multi-agent systems decompose complex tasks into interacting roles, but most remain manually orchestrated by prompts, tools, and control rules

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Variational Inference for Evidential Deep Learning

DGX agent

arXiv:2605.26477v1 Announce Type: new Abstract: While Deep Neural Networks (DNNs) achieve remarkable performance, their tendency to produce overconfident predictions. Evidential Deep Learning (EDL) mi

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Vectors Are Not Neutral: Sensitive-Information Inference from Exported LLM Representations in Summarization

DGX agent

arXiv:2605.26433v1 Announce Type: new Abstract: Large language model (LLM) summarization systems may pass compact vector representations of private inputs to downstream retrieval, monitoring, audit, o

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Verus-SpecGym: An Agentic Environment for Evaluating Specification Autoformalization

DGX agent

arXiv:2605.26457v1 Announce Type: cross Abstract: AI coding agents are increasingly used to write real-world software, but ensuring that their outputs are correct remains a fundamental challenge. Form

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

VISTA: An End-to-End Benchmark for Visual Spec-to-Web-App Coding Agents

DGX agent

arXiv:2605.26144v1 Announce Type: cross Abstract: We present VISTA (VIsual Spec-To-App Benchmark), a benchmark for evaluating the end-to-end web-app generation capabilities of LLM-based agents. Unlike

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

VisualNeedle: Benchmarking Active Visual Search in Information-Dense Scenes

DGX agent

arXiv:2605.26380v1 Announce Type: cross Abstract: Frontier multimodal large language models (MLLMs) have been reported to achieve over 90% accuracy on fine-grained perception benchmarks. However, such

model-releasesarxiv-cs-ai
27 May 2026
← Previous
1…263264265266267…475
Next →