AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,566 results
30 Jun 2026

Internalized Reasoning for Long-Context Visual Document Understanding

Model ReleasesDGX agent

arXiv:2604.02371v2 Announce Type: replace-cross Abstract: Visual long-document understanding is critical for enterprise, legal, and scientific applications, yet the best performing open recipes have n

Introducing Claude Sonnet 5 on AWS: Anthropic’s most capable Sonnet model

Model ReleasesDGX agent

Today, we’re excited to announce the availability of Anthropic’s most advanced Sonnet model, Claude Sonnet 5, on Amazon Bedrock and Claude Platform on AWS. Claude Sonnet 5 is the first Sonnet model of

Introducing GeneBench-Pro

Model ReleasesDGX agent

GeneBench-Pro is a comprehensive benchmarking tool or dataset introduced by OpenAI designed to evaluate AI model performance on genomic and gene-related tasks. It likely provides standardized metrics


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Introducing SWE-Together: a multi-turn benchmark built from real user–agent coding sessions. Coding agents are often benchmarked like exam-t…

Model ReleasesDGX agent

Introducing SWE-Together: a multi-turn benchmark built from real user–agent coding sessions. Coding agents are often benchmarked like exam-takers: given the full spec up front, then graded on the fina

It Lied to a Doctor to Buy Poison Ingredients: Quantifying Real-World Misuse of Phone-use Agents

Model ReleasesDGX agent

arXiv:2606.27944v1 Announce Type: cross Abstract: Phone-use Agents can execute complex tasks end to end across real mobile applications. By operating a real device on the user's behalf, they reach far

JuZhou 1.0 Technical Report: The First Edge-Native Text-to-Image Foundation Model Trained Entirely on China-Developed AI Accelerators

Model ReleasesDGX agent

arXiv:2606.28421v1 Announce Type: cross Abstract: Text-to-image (T2I) diffusion models typically require substantial computational resources and cloud infrastructure, posing significant challenges for

Know Before You Fetch: Calibrated Retrieval-Budget Allocation for Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2606.29959v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) typically retrieves a fixed number of passages for every query. This is wasteful when the reader already knows th

KnowsTFM: Knowledge-Informed Fine-Tuning of Small Tabular Foundation Models

Model ReleasesDGX agent

arXiv:2606.30258v1 Announce Type: cross Abstract: Tabular foundation models have advanced deep learning for tabular data by delivering strong default performance across many small and medium tasks. Ye

KrishokChat: A Citation-Grounded Dataset and Benchmark for Bengali Agricultural Advisory

Model ReleasesDGX agent

arXiv:2606.29243v1 Announce Type: new Abstract: We present KrishokChat, the first citation-grounded Bengali agricultural instruction-tuning dataset for crop advisory in low-resource settings. We estab

Labeling Training Data for Entity Matching Using Large Language Models

Model ReleasesDGX agent

arXiv:2606.28823v1 Announce Type: new Abstract: Recent large language models (LLMs) achieve strong performance on entity matching without requiring task-specific training data. However, applying these

LatSearch: Latent Reward-Guided Search for Faster Inference-Time Scaling in Video Diffusion

Model ReleasesDGX agent

arXiv:2603.14526v2 Announce Type: replace Abstract: The recent success of inference-time scaling in large language models has inspired similar explorations in video diffusion. In particular, motivated

LaVPR: Benchmarking Language and Vision for Place Recognition

Model ReleasesDGX agent

arXiv:2602.03253v2 Announce Type: replace Abstract: Visual Place Recognition (VPR) often fails under extreme environmental changes and perceptual aliasing. Beyond these limitations, standard systems c

Learning Cross-view Correspondences for Geo-localization on Planetary Surfaces

Model ReleasesDGX agent

arXiv:2606.29821v1 Announce Type: new Abstract: Maintaining global position awareness is a fundamental challenge for planetary surface exploration, since satellite-based positioning systems are unavai

Learning from Reliable Latent Prompts for Visual Recognition with Missing Modalities

Model ReleasesDGX agent

arXiv:2606.30597v1 Announce Type: new Abstract: Large-scale multimodal models (LMMs) have achieved superior performance in visual recognition by synergizing information across diverse, massive-scale p

Learning from samples: inverse problems over measures

Model ReleasesDGX agent

arXiv:2505.07124v3 Announce Type: replace Abstract: We study inverse problems where an unknown potential is observed only through samples from the measure it induces by a convex variational principle.

Learning to Bid in Discriminatory Auctions with Budget Constraints

Model ReleasesDGX agent

arXiv:2606.29252v1 Announce Type: new Abstract: We study repeated bidding in multi-unit discriminatory (pay-as-bid) auctions for a single bidder with per-round utility equal to value minus alpha times

Learning to Route and Schedule LLMs from User Retrials via Contextual Queueing Bandits

Model ReleasesDGX agent

arXiv:2602.02061v2 Announce Type: replace Abstract: Explosive demands for LLMs often cause user queries to accumulate in server queues, requiring efficient routing (query-LLM matching) and scheduling

Learning to Segment Liquids in Real-world Images

Model ReleasesDGX agent

arXiv:2601.00940v2 Announce Type: replace Abstract: Liquids like water, wine and medicine are everywhere. However, limited attention has been given to the task of segmenting liquids, hindering the abi

LEDGER: Scaling Agentic Document Editing with Dependency-aware Graph Retrieval

Model ReleasesDGX agent

arXiv:2606.28379v1 Announce Type: cross Abstract: We introduce LEDGER to tackle the novel context engineering challenge of agentic document editing, where localized edits to long, structured documents

LEIQ-Assessor: Multi-dimensional Quality Assessment of Low-light Enhanced Images via Multi-task Learning

Model ReleasesDGX agent

arXiv:2606.29752v1 Announce Type: new Abstract: Low-light image enhancement algorithms (LIEAs) aim to improve the visibility of images captured under poor illumination. However, the enhancement proces

Little Brains, Big Feats: Exploring Compact Language Models

Model ReleasesDGX agent

arXiv:2606.30062v1 Announce Type: cross Abstract: While large language models have been dominating the research landscape recently, small language models remain highly relevant across various domains;

LLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via a Proprioceptive Dashboard

Model ReleasesDGX agent

arXiv:2606.30005v1 Announce Type: new Abstract: Long-horizon tool agents are bottlenecked by how their context grows toward the limits of the context window. Recent systems make context management age

LLM-based Multimodal Personality Recognition via Facial Action Unit-Text Semantic Fusion

Model ReleasesDGX agent

arXiv:2606.29900v1 Announce Type: cross Abstract: Personality recognition in asynchronous video interviews (AVIs) has become increasingly important due to their widespread adoption in modern recruitme

LLM-Guided Planning for Multi-hop Reasoning over Multimodal Nuclear Regulatory Documents

Model ReleasesDGX agent

arXiv:2606.29399v1 Announce Type: new Abstract: Reviewing nuclear regulatory documents requires multi-hop reasoning across tens of thousands of pages, where judgments depend on evidence assembled acro

LoGSAM: Parameter-Efficient Cross-Modal Grounding for MRI Segmentation

Model ReleasesDGX agent

arXiv:2603.17576v3 Announce Type: replace Abstract: Precise localization and delineation of brain tumors using magnetic resonance imaging (MRI) are essential for planning therapy and guiding surgical

Longitudinal Lesion Inpainting in Brain MRI via 3D Region Aware Diffusion

Model ReleasesDGX agent

arXiv:2603.05693v2 Announce Type: replace-cross Abstract: Accurate longitudinal analysis of brain MRI is often hindered by evolving lesions, which bias automated neuroimaging pipelines. While deep gen

“Loop engineering” is a hot buzzphrase after mentions of it by Boris Cherny (Claude Code’s creator) and Peter Steinberger (OpenClaw's creato…

Model ReleasesDGX agent

“Loop engineering” is a hot buzzphrase after mentions of it by Boris Cherny (Claude Code’s creator) and Peter Steinberger (OpenClaw's creator) went viral on social media. Loops are now a key part of h

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models

Model ReleasesDGX agent

arXiv:2601.05366v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed as agents that invoke external tools through structured function calls. While recent wo

Love how Google continues to drive down the cost of building with their models. <4s image and $0.034 / 1K image. Wow! We have a bunch of stu…

Model ReleasesDGX agent

Love how Google continues to drive down the cost of building with their models. <4s image and 0.034 / 1K image. Wow! We have a bunch of stuff (education & research) we're building @dair_ai using Nano

LUMEN: Cost-Transparent Multi-Agent Pipeline for Automated Systematic Review and Meta-Analysis

Model ReleasesDGX agent

arXiv:2606.28362v1 Announce Type: cross Abstract: Systematic reviews and meta-analyses (SR/MA) remain the gold standard for evidence synthesis, yet completing one typically requires 67 weeks and subst

LWDrive: Layer-Wise World-Model-Guided Vision-Language Model Planning for Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.29879v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) provide powerful semantic understanding and commonsense reasoning for End-to-End Autonomous Driving (E2E-AD) planning. H

MACROCAST: A Vintage-Consistent Time Series Foundation Model for Real-Time Macroeconomic Forecasting

Model ReleasesDGX agent

arXiv:2606.28670v1 Announce Type: cross Abstract: We introduce MACROCAST, a lightweight Time Series Foundation Model (TSFM) for real-time macroeconomic forecasting. Existing TSFMs suffer from data lea

MaDI-Bench: An End-to-End Data Integration Benchmark

Model ReleasesDGX agent

arXiv:2606.30371v1 Announce Type: cross Abstract: Data integration combines heterogeneous data sets into a single, coherent representation. Data integration involves a sequence of interdependent tasks

Making Multimodal LLMs Reliable Chart Data Extractors: A Benchmark and Training Framework

Model ReleasesDGX agent

arXiv:2606.29808v1 Announce Type: cross Abstract: Chart data extraction, which reverse-engineers data tables from chart images, is essential for reproducibility, analysis, retrieval, and redesign. Exi

MAM-AI: An On-Device Medical Retrieval-Augmented Generation System for Nurses and Midwives in Zanzibar

Model ReleasesDGX agent

arXiv:2606.29580v1 Announce Type: new Abstract: Maternal and newborn mortality remain among the highest in sub-Saharan Africa, where midwifery care is often delivered by nurses who lack midwifery trai

mamabench and mamaretrieval: Benchmarks for Evaluating Medical Retrieval-Augmented Generation in Maternal, Neonatal, and Reproductive Health

Model ReleasesDGX agent

arXiv:2606.29467v1 Announce Type: new Abstract: Medical question-answering benchmarks rarely cover the maternal, neonatal, child, and reproductive-health questions a nurse-midwife asks, and, to our kn

Managing the Human Fallback: Skill Investment Under Improving AI and Worker Mobility

Model ReleasesDGX agent

arXiv:2606.29111v1 Announce Type: new Abstract: When firms deploy autonomous AI, they must decide how much work to leave to the system and how much to keep workers engaged. This decision affects curre

MCP Server Architecture Patterns for LLM-Integrated Applications

Model ReleasesDGX agent

arXiv:2606.30317v1 Announce Type: cross Abstract: The Model Context Protocol (MCP), introduced by Anthropic in November 2024, defines a standardized interface for connecting large language models (LLM

Meituan open-sources LongCat-2.0, a 1.6T-parameter model that it says was trained on a 50K-chip cluster of domestic Chinese processors, without giving details (Reuters)

Model ReleasesDGX agent

Reuters: Meituan open-sources LongCat-2.0, a 1.6T-parameter model that it says was trained on a 50K-chip cluster of domestic Chinese processors, without giving details — China's food delivery giant Me

MemDelta: Controlled Baselines and Hidden Confounds in Agent Memory Evaluation

Model ReleasesDGX agent

arXiv:2606.29914v1 Announce Type: new Abstract: Agent memory systems are increasingly evaluated against RAG and full-context baselines, but reported gains often mix changes in the memory method with c

MemLeak: Diagnosing Information Leaks in Multimodal Agent Memory

Model ReleasesDGX agent

arXiv:2606.29788v1 Announce Type: new Abstract: When a multimodal AI agent is asked to forget a fact, current memory systems usually delete the text entry and report success. We find that the fact can

Memory-Managed Long-Context Attention: A Preliminary Study of Editable Request-Local Memory

Model ReleasesDGX agent

arXiv:2606.28876v1 Announce Type: new Abstract: Long-context language models often conflate two different goals: compressing history into an efficient state, and maintaining reliable long-term memory.

MESA: Prioritizing Vulnerable Communication Channels for Securing Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2606.30602v1 Announce Type: cross Abstract: Multi-agent systems (MAS) are increasingly used to automate complex, distributed workflows. However, their inter-agent communication channels introduc

meta-pipe: An LLM-agent pipeline for end-to-end automated systematic review and meta-analysis

Model ReleasesDGX agent

arXiv:2606.28363v1 Announce Type: cross Abstract: Objective: To describe the architecture and design rationale of meta-pipe, an open-source large language model (LLM)-agent pipeline that integrates th

MirrorCode: AI can rebuild entire programs from behavior alone

Model ReleasesDGX agent

arXiv:2606.30182v1 Announce Type: new Abstract: AI models are rapidly improving at autonomous coding, as shown by benchmark progress and one-off demonstrations such as AI implementing a C compiler. Ho

MixSarc: A Bangla-English Code-Mixed Corpus for Implicit Meaning Identification

Model ReleasesDGX agent

arXiv:2602.21608v2 Announce Type: replace Abstract: Bangla-English code-mixing is widespread across South Asian social media, yet resources for implicit meaning identification in this setting remain s

Model Merging to Evolution: Parameter Space Exploration for Expert Models

Model ReleasesDGX agent

arXiv:2606.28373v1 Announce Type: cross Abstract: Model merging integrates the capabilities of multiple expert models to create strong models for multiple tasks without additional training, thereby re

Model Predictive Current Control with Harmonic Correction for Single-Phase AC-DC EV Charging

Model ReleasesDGX agent

arXiv:2606.30397v1 Announce Type: cross Abstract: The increasing integration of Electric Vehicles (EVs) has imposed a growing harmonic challenge on the power grid. For AC/DC Power Factor Correction (P

Monte Carlo Energy Aggregation for Mobile 3D Gaussian Splatting

Model ReleasesDGX agent

arXiv:2606.30017v1 Announce Type: new Abstract: Recent advances in 3D Gaussian Splatting have demonstrated unprecedented success in novel view synthesis. However, the substantial inference and storage

Morphing into Hybrid Attention Models

Model ReleasesDGX agent

arXiv:2606.30562v1 Announce Type: new Abstract: Hybrid attention models improve long-context efficiency by retaining only a subset of full-attention layers and replacing the remaining layers with line

MotionAtlas: Detailed Region Captioning for Motion-Centric Videos

Model ReleasesDGX agent

arXiv:2606.29531v1 Announce Type: cross Abstract: We propose MotionAtlas, a system for detailed captioning of motion-centric videos, comprising (1) a dedicated human-annotated benchmark, (2) a scalabl

MSA-UNet3+: Multi-Scale Attention UNet3+ with New Supervised Prototypical Contrastive Loss for Coronary DSA Image Segmentation

Model ReleasesDGX agent

arXiv:2504.05184v4 Announce Type: replace-cross Abstract: Accurate segmentation of coronary Digital Subtraction Angiography (DSA) images is essential for diagnosing and treating coronary artery diseas

Multi-Agent Route Planning as a QUBO Problem

Model ReleasesDGX agent

arXiv:2602.07913v2 Announce Type: replace Abstract: Multi-Agent Route Planning considers selecting vehicles, each associated with a single predefined route, such that route-level coverage utility is m

Multi-Agent Routing as Set-Valued Prediction: A WildChat Benchmark and Cost-Aware Evaluation

Model ReleasesDGX agent

arXiv:2606.28925v1 Announce Type: cross Abstract: Tool and agent routing from natural-language prompts is naturally a set-valued prediction problem: a single query may require multiple agents, while o

Multi-Agentic System Leveraging Open-Source LLMs to Mitigate Disinformation Threats

Model ReleasesDGX agent

arXiv:2606.30259v1 Announce Type: new Abstract: In contemporary societies, the threat of disinformation has reached alarming levels, exacerbated by the proliferation of electronic communication, socia

Multi-scale Object-Aware Gaze Estimation via Geometric Reasoning

Model ReleasesDGX agent

arXiv:2606.29334v1 Announce Type: new Abstract: Gaze target estimation aims to predict the semantic object an observer fixates upon within an image, a task deeply rooted in the object-oriented nature

Multimodal Graph RAG for Long-range Visually Rich Document Understanding

Model ReleasesDGX agent

arXiv:2606.28780v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are widely applied to visual document understanding. However, comprehending long documents remains an issue b

Multimodal Mathematical Reasoning with Diverse Solving Perspective

Model ReleasesDGX agent

arXiv:2507.02804v2 Announce Type: replace Abstract: Recent progress in large-scale reinforcement learning (RL) has notably enhanced the reasoning capabilities of large language models (LLMs), especial

Multiply Robust Causal Mediation Analysis with Continuous Treatments

Model ReleasesDGX agent

arXiv:2105.09254v4 Announce Type: replace-cross Abstract: In many applications, researchers are interested in the direct and indirect causal effects of a treatment or exposure on an outcome of interes

Muon learns balanced solutions in matrix factorization without slow saddle-to-saddle dynamics

Model ReleasesDGX agent

arXiv:2606.30509v1 Announce Type: new Abstract: Matrix factorization (i.e., problems of the form min_{mathbf{P},mathbf{Q}} |mathbf{M}^star - mathbf{P}^opmathbf{Q}|_F^2) is a minimal learning problem t

← Previous
1…119120121122123…377
Next →