AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
15 May 2026

AMiD: Knowledge Distillation for LLMs with alpha-mixture Assistant Distribution

SafetyDGX agent

arXiv:2510.15982v3 Announce Type: replace-cross Abstract: Autoregressive large language models (LLMs) have achieved remarkable improvement across many tasks but incur high computational and memory cos

APWA: A Distributed Architecture for Parallelizable Agentic Workflows

AgentsDGX agent

arXiv:2605.15132v1 Announce Type: new Abstract: Autonomous multi-agent systems based on large language models (LLMs) have demonstrated remarkable abilities in independently solving complex tasks in a

AVEX: What Matters for Animal Vocalization Encoding

ResearchDGX agent

arXiv:2508.11845v3 Announce Type: replace-cross Abstract: Bioacoustics, the study of sounds produced by living organisms, plays a vital role in conservation, biodiversity monitoring, and behavioral st

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

b9159

Local AiDGX agent

b9159 is a release of llama.cpp published on May 14, 2026 . llama.cpp is a C/C++ implementation of large language model inference that enables efficient LLM execution on consumer hardware with minimal

b9169

Local AiDGX agent

Release b9169 of llama.cpp includes updates to multi-token multimodal decoding (mtmd) functionality, adding chunks and fixing preprocessing for Qwen3A models . The changes include attention mask imple

Bad Seeing or Bad Thinking? Rewarding Perception for Vision-Language Reasoning

AgentsDGX agent

arXiv:2605.14054v1 Announce Type: new Abstract: Achieving robust perception-reasoning synergy is a central goal for advanced Vision-Language Models (VLMs). Recent advancements have pursued this goal v

Boosting Reinforcement Learning with Verifiable Rewards via Randomly Selected Few-Shot Guidance

SafetyDGX agent

arXiv:2605.15012v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has achieved great success in developing Large Language Models (LLMs) with chain-of-thought roll

Bridging the Rural Healthcare Gap: A Cascaded Edge-Cloud Architecture for Automated Retinal Screening

Local AiDGX agent

arXiv:2605.14108v1 Announce Type: cross Abstract: Diabetic Retinopathy (DR) is one of the leading causes of preventable blindness, yet rural regions often lack the specialists and infrastructure neede

Characterizing the visual representation of objects from the child's view

TutorialsDGX agent

arXiv:2605.14990v1 Announce Type: new Abstract: Children acquire object category representations from their everyday experiences in the first few years of life. What do the inputs to this learning pro

CHASM: Cross-frequency Harmonized Axis-Separable Mixing for Spectral Token Operators

SafetyDGX agent

arXiv:2605.14727v1 Announce Type: new Abstract: Spectral token mixers based on Fourier transforms provide an efficient way to model global interactions in visual feature maps. Existing designs often e

ChromaFlow: A Negative Ablation Study of Orchestration Overhead in Tool-Augmented Agent Evaluation

AgentsDGX agent

arXiv:2605.14102v1 Announce Type: new Abstract: Autonomous language-model agents increasingly combine planning, tool use, document processing, browsing, code execution, and verification loops. These c

CoRDS: Coreset-based Representative and Diverse Selection for Streaming Video Understanding

ResearchDGX agent

arXiv:2605.14310v1 Announce Type: new Abstract: Streaming video understanding with large vision-language models (VLMs) requires a compact memory that can support future reasoning over an ever-growing

cred to @thatguybg, playing with micro yesterday inspired this (consider this my feature wish list)

AgentsDGX agent

This post shares feature requests and ideas inspired by experimentation with micro (likely a small language model or tool), credited to another developer. It appears to be a wishlist of desired capabi

CSI-JEPA: Towards Foundation Representations for Ubiquitous Sensing with Minimal Supervision

ApplicationsDGX agent

arXiv:2605.14171v1 Announce Type: new Abstract: Channel state information (CSI) provides a widely available sensing modality for human and environment perception, but existing CSI sensing models usual

Deepchecks: Evaluating Retrieval-Augmented Generation (RAG)

SafetyDGX agent

arXiv:2605.14488v1 Announce Type: new Abstract: Large Language Models (LLMs) augmented with Retrieval-Augmented Generation (RAG) techniques are revolutionizing applications across multiple domains, su

DermAgent: A Self-Reflective Agentic System for Dermatological Image Analysis with Multi-Tool Reasoning and Traceable Decision-Making

AgentsDGX agent

arXiv:2605.14403v1 Announce Type: new Abstract: Dermatological diagnosis requires integrating fine-grained visual perception with expert clinical knowledge. Although Multimodal Large Language Models (

Do We Really Need External Tools to Mitigate Hallucinations? SIRA: Shared-Prefix Internal Reconstruction of Attribution

ResearchDGX agent

arXiv:2605.14621v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) often hallucinate when language priors dominate weak or ambiguous visual evidence. Existing contrastive decoding

DSSP: Diffusion State Space Policy with Full-History Encoding

SafetyDGX agent

arXiv:2605.14598v1 Announce Type: new Abstract: Diffusion-based imitation learning has shown strong promise for robot manipulation. However, most existing policies condition only on the current observ

Dual-Dimensional Consistency: Balancing Budget and Quality in Adaptive Inference-Time Scaling

ResearchDGX agent

arXiv:2605.15100v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable abilities in reasoning. However, maximizing their potential through inference-time scaling fac

Dynamic Latent Routing

ResearchDGX agent

arXiv:2605.14323v1 Announce Type: cross Abstract: We investigate the temporal concatenation of sub-policies in Markov Decision Processes (MDP) with time-varying reward functions. We introduce General

Dynamics of the Transformer Residual Stream: Coupling Spectral Geometry to Network Topology

ApplicationsDGX agent

arXiv:2605.14258v1 Announce Type: cross Abstract: Large language models are remarkably capable, yet how computation propagates through their layers remains poorly understood. A growing line of work tr

EBMs are so back! Aleph now leads the major formal reasoning benchmarks.

ResearchDGX agent

Energy-Based Models (EBMs) have demonstrated renewed competitiveness in formal reasoning tasks, with Aleph achieving leading performance on major benchmarks in this domain. This represents a significa

Efficient Generative Retrieval for E-commerce Search with Semantic Cluster IDs and Expert-Guided RL

SafetyDGX agent

arXiv:2605.14434v1 Announce Type: cross Abstract: Generative retrieval offers a promising alternative by unifying the fragmented multi-stage retrieval process into a single end-to-end model. However,

Efficient Multi-objective Prompt Optimization via Pure-exploration Bandits

ResearchDGX agent

arXiv:2605.14553v1 Announce Type: cross Abstract: Prompt engineering has become central to eliciting the capabilities of large language models (LLMs). At its core lies prompt selection -- efficiently

Evidential Reasoning Advances Interpretable Real-World Disease Screening

ApplicationsDGX agent

arXiv:2605.15171v1 Announce Type: cross Abstract: Disease screening is critical for early detection and timely intervention in clinical practice. However, most current screening models for medical ima

EXCLUSIVE: Runway is one of the most interesting inflection point stories in AI right now. Started as a tool for filmmakers. Now going after…

IndustryDGX agent

EXCLUSIVE: Runway is one of the most interesting inflection point stories in AI right now. Started as a tool for filmmakers. Now going after world models/ universe simulation, and by extension, everyt

Explainable Detection of Depression Status Shifts from User Digital Traces

ResearchDGX agent

arXiv:2605.14995v1 Announce Type: new Abstract: Every day, users generate digital traces (e.g., social media posts, chats, and online interactions) that are inherently timestamped and may reflect aspe

Exploitation of Hidden Context in Dynamic Movement Forecasting: A Neural Network Journey from Recurrent to Graph Neural Networks and General Purpose Transformers

ResearchDGX agent

arXiv:2605.14855v1 Announce Type: cross Abstract: Forecasting within signal processing pipelines is crucial for mitigating delays, particularly in predicting the dynamic movements of objects such as N

FactNet: A Billion-Scale Knowledge Graph for Multilingual Factual Grounding

ResearchDGX agent

arXiv:2602.03417v2 Announce Type: replace Abstract: Large language models hallucinate factual claims and struggle to ground their outputs in retrievable evidence, particularly in non-English languages

Feature Visualization Recovers Known Cortical Selectivity from TRIBE v2

ResearchDGX agent

arXiv:2605.13904v1 Announce Type: cross Abstract: Brain encoder models predict cortical fMRI responses from the internal activations of pretrained vision and language networks, and are typically evalu

From Plans to Pixels: Learning to Plan and Orchestrate for Open-Ended Image Editing

AgentsDGX agent

arXiv:2605.15181v1 Announce Type: new Abstract: Modern image editing models produce realistic results but struggle with abstract, multi step instructions (e.g., ``make this advertisement more vegetari

Functional-level Uncertainty Quantification for Calibrated Fine-tuning on LLMs

ResearchDGX agent

arXiv:2410.06431v5 Announce Type: replace Abstract: Accurate uncertainty quantification in large language models (LLMs) is essential for reliable confidence estimation, yet fine-tuned LLMs often becom

Generative Floor Plan Design with LLMs via Reinforcement Learning with Verifiable Rewards

ResearchDGX agent

arXiv:2605.14117v1 Announce Type: cross Abstract: An AI system for professional floor plan design must precisely control room dimensions and areas while respecting the desired connectivity between roo

GeoViSTA: Geospatial Vision-Tabular Transformer for Multimodal Environment Representation

Local AiDGX agent

arXiv:2605.14406v1 Announce Type: cross Abstract: Large-scale pretraining on Earth observation imagery has yielded powerful representations of the natural and built environment. However, most existing

Google updates its spam rules to include attempts to ‘manipulate’ AI

SafetyDGX agent

Google updated its spam policy to mark attempts to 'manipulate' its AI model in search results as spam, including results in AI Overview or AI Mode in Search, as Search Engine Land reports: 'In the co

Graphs of Research: Citation Evolution Graphs as Supervision for Research Idea Generation

ResearchDGX agent

arXiv:2605.14790v1 Announce Type: cross Abstract: Research idea generation is the innovation-driving step of automated scientific research. Recently, large language models (LLMs) have shown potential

Hand-in-the-Loop: Improving Dexterous VLA via Seamless Interventional Correction

SafetyDGX agent

arXiv:2605.15157v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are prone to compounding errors in dexterous manipulation, where high-dimensional action spaces and contact-rich d

Head Forcing: Long Autoregressive Video Generation via Head Heterogeneity

ResearchDGX agent

arXiv:2605.14487v1 Announce Type: cross Abstract: Autoregressive video diffusion models support real-time synthesis but suffer from error accumulation and context loss over long horizons. We discover

How to Scale Mixture-of-Experts: From muP to the Maximally Scale-Stable Parameterization

TutorialsDGX agent

arXiv:2605.14200v1 Announce Type: new Abstract: Recent frontier large language models predominantly rely on Mixture-of-Experts (MoE) architectures. Despite empirical progress, there is still no princi

ICYMI: 1️⃣ LangSmith Engine 2️⃣ SmithDB 3️⃣ Managed Deep Agents 4️⃣ LangSmith Sandboxes: Now Generally Available 5️⃣ Context Hub 6️⃣ LangSmi…

AgentsDGX agent

ICYMI: 1️⃣ LangSmith Engine 2️⃣ SmithDB 3️⃣ Managed Deep Agents 4️⃣ LangSmith Sandboxes: Now Generally Available 5️⃣ Context Hub 6️⃣ LangSmith LLM Gateway 7️⃣ Sandboxes, Prebuilt agents, + free model

IsoNet: Spatially-aware audio-visual target speech extraction in complex acoustic environments

ResearchDGX agent

arXiv:2605.14736v1 Announce Type: cross Abstract: Target speech extraction remains difficult for compact devices because monaural neural models lack spatial evidence and classical beamformers lose res

Lang2MLIP: End-to-End Language-to-Machine Learning Interatomic Potential Development with Autonomous Agentic Workflows

AgentsDGX agent

arXiv:2605.14527v1 Announce Type: new Abstract: Developing machine learning interatomic potentials (MLIPs) for complex materials systems remains challenging because it requires expertise in atomistic

@LangChain’s Interrupt 2026 conference was a blast!! such a pleasure to with @VictorMoreira16 about deep agents! ICYMI: we just dropped v0.6…

AgentsDGX agent

@LangChain’s Interrupt 2026 conference was a blast!! such a pleasure to with @VictorMoreira16 about deep agents! ICYMI: we just dropped v0.6, which is focused on performance at the model, harness, and

Language Generation as Optimal Control: Closed-Loop Diffusion in Latent Control Space

SafetyDGX agent

arXiv:2605.14531v1 Announce Type: new Abstract: This work reformulates language generation as a stochastic optimal control problem, providing a unified theoretical perspective to analyze autoregressiv

Language-Induced Priors for Domain Adaptation

TutorialsDGX agent

arXiv:2605.14301v1 Announce Type: new Abstract: Domain adaptation faces a fundamental paradox in the cold-start regime. When target data is scarce, statistical methods fail to distinguish relevant sou

LATERN: Test-Time Context-Aware Explainable Video Anomaly Detection

SafetyDGX agent

arXiv:2605.15054v1 Announce Type: new Abstract: Vision-language models (VLMs) have recently emerged as a promising paradigm for video anomaly detection (VAD) due to their strong visual reasoning abili

LEMON: Learning Executable Multi-Agent Orchestration via Counterfactual Reinforcement Learning

Local AiDGX agent

arXiv:2605.14483v1 Announce Type: new Abstract: Large language models (LLMs) have become a strong foundation for multi-agent systems, but their effectiveness depends heavily on orchestration design. A

MAPLE: Self-Supervised Learning-Enhanced Nonlinear Dimensionality Reduction for Visual Analysis

ResearchDGX agent

arXiv:2601.20173v2 Announce Type: replace Abstract: We present a new nonlinear dimensionality reduction method, MAPLE, that enhances UMAP by improving manifold modeling. MAPLE employs a self-supervise

Masked Autoencoders with Limited Data: Does It Work? A Fine-Grained Bioacoustics Case Study

ApplicationsDGX agent

arXiv:2605.14031v1 Announce Type: cross Abstract: Bioacoustic recognition requires fine-grained acoustic understanding to distinguish similar-sounding species. However, many large-scale data repositor

Matrix-Space Reinforcement Learning for Reusing Local Transition Geometry

Local AiDGX agent

arXiv:2605.14304v1 Announce Type: cross Abstract: Compositional generalization in sequential decision-making requires identifying which parts of prior rollouts remain useful for new tasks. Existing me

Medical Report Generation: A Hierarchical Task Structure-Based Cross-Modal Causal Intervention Framework

SafetyDGX agent

arXiv:2511.02271v2 Announce Type: replace Abstract: Medical Report Generation (MRG) is a key part of modern medical diagnostics, as it automatically generates reports from radiological images to reduc

MemLineage: Lineage-Guided Enforcement for LLM Agent Memory

AgentsDGX agent

arXiv:2605.14421v1 Announce Type: cross Abstract: We introduce MemLineage, a defense for LLM agent memory that attaches both cryptographic provenance and LLM-mediated derivation lineage to every entry

MetaMoE: Diversity-Aware Proxy Selection for Privacy-Preserving Mixture-of-Experts Unification

ResearchDGX agent

arXiv:2605.14289v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models scale capacity by combining specialized experts, but most existing approaches assume centralized access to training da

MHSA: A Lightweight Framework for Mitigating Hallucinations via Steered Attention in LVLMs

ResearchDGX agent

arXiv:2605.14966v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) have achieved remarkable performance across diverse multimodal tasks, yet they continue to suffer from hallucinat

MiVE: Multiscale Vision-language features for reference-guided video Editing

ResearchDGX agent

arXiv:2605.14664v1 Announce Type: new Abstract: Reference-guided video editing takes a source video, a text instruction, and a reference image as inputs, requiring the model to faithfully apply the in

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning

TutorialsDGX agent

arXiv:2509.22746v2 Announce Type: replace Abstract: Current visual reasoning methods mainly focus on exploring specific reasoning modes. Although improvements can be achieved in particular domains, th

Monitoring Data-aware Temporal Properties (Extended Version)

ResearchDGX agent

arXiv:2605.14666v1 Announce Type: new Abstract: Dynamic systems in AI are often complex and heterogeneous, so that an internal specification is not accessible and verification techniques such as model

OmniDrop: Layer-wise Token Pruning for Omni-modal LLMs via Query-Guidance

ResearchDGX agent

arXiv:2605.14458v1 Announce Type: new Abstract: Omni-modal large language models have demonstrated remarkable potential in holistic multimodal understanding; however, the token explosion caused by hig

Optimal Pattern Detection Tree for Symbolic Rule-Based Classification

ApplicationsDGX agent

arXiv:2605.14374v1 Announce Type: cross Abstract: Pattern discovery in data plays a crucial role across diverse domains, including healthcare, risk assessment, and machinery maintenance. In contrast t

Paper of the day! https://huggingface.co/papers/2605.13301

IndustryDGX agent

Paper of the day! https://huggingface.co/papers/2605.13301 We’re releasing a 30B-A3B reasoning model that reaches gold-medal level across both physics and math Olympiad evaluations: IPhO directly, and

← Previous
1…773774775776777…1017
Next →