AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlog
90,338Total entries
1Added by human
90,337Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,187 results
Safety

Is Diversity All You Need for Scalable Robotic Manipulation?

DGX agent

arXiv:2507.06219v2 Announce Type: replace Abstract: Data scaling has driven remarkable success in foundation models for Natural Language Processing (NLP) and Computer Vision (CV), yet the principles o

safetyarxiv-cs-ro
5 Jun 2026
Applications
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

LANTERN: Layered Archival and Temporal Episodic Retrieval Network for Long-Context LLM Conversations

DGX agent

arXiv:2606.05182v1 Announce Type: new Abstract: Large language models discard critical details when conversation history is compacted to fit within finite context windows. We present LANTERN (Layered

applicationsarxiv-cs-cl
5 Jun 2026
Model Releases

LLM-Guided ANN Index Optimization for Human-Object Interaction Retrieval

DGX agent

arXiv:2606.05489v1 Announce Type: new Abstract: Retrieval systems underpin modern AI applications -- spanning visual search, recommendation engines, and multi-modal question answering. Modern multi-st

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

MAviS: A Multimodal Conversational Assistant For Avian Species

DGX agent

arXiv:2603.07294v2 Announce Type: replace Abstract: Fine-grained understanding and species-specific multimodal question answering are vital for advancing biodiversity conservation and ecological monit

model-releasesarxiv-cs-cv
5 Jun 2026
Local Ai

Mechanistic Insights into Functional Sparsity in Multimodal LLMs via CoRe Heads

DGX agent

arXiv:2606.05843v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) demonstrate remarkable proficiency on complex vision-language tasks, the mechanisms by which they extract

local-aiarxiv-cs-cl
5 Jun 2026
Applications

OneReason Technical Report

DGX agent

arXiv:2606.06260v1 Announce Type: cross Abstract: Generative recommendation models in the OneRec family have been widely deployed in many real-world services, such as short-video, live-streaming, adve

applicationsarxiv-cs-cl
5 Jun 2026
Research

PAR3D: A Unified 3D-MLLM with Part-Aware Representation for Scene Understanding

DGX agent

arXiv:2606.06485v1 Announce Type: new Abstract: Recent advances in 3D multimodal large language models (3D-MLLMs) have enabled unified solutions for 3D scene understanding tasks, including visual ques

researcharxiv-cs-cv
5 Jun 2026
Research

Physics in 2-Steps: Locking Motion Priors Before Visual Refinement Erases Them

DGX agent

arXiv:2606.06361v1 Announce Type: new Abstract: Image-to-Video diffusion models leverage input images to generate visually stunning content, yet frequently produce motion that violates physical laws.

researcharxiv-cs-cv
5 Jun 2026
Research

Predictable Scaling Laws of Optimal Hyperparameters for LLM Continued Pre-training

DGX agent

arXiv:2606.05610v1 Announce Type: new Abstract: The efficacy of continued pre-training for Large Language Models (LLMs) hinges upon hyperparameter configurations, such as learning rate and batch size.

researcharxiv-cs-cl
5 Jun 2026
Local Ai

ProSarc: Prosody-Aware Sarcasm Recognition Framework via Temporal Prosodic Incongruity

DGX agent

arXiv:2606.06168v1 Announce Type: cross Abstract: We present ProSarc, an audio-only framework that detects sarcasm by modelling temporal prosodic incongruity, that is, the mismatch between local proso

local-aiarxiv-cs-cl
5 Jun 2026
Model Releases

Representing Research Attention as Contextually Structured Flows

DGX agent

arXiv:2606.05895v1 Announce Type: new Abstract: Research attention is widely used as an indicator of visibility, influence, and societal uptake, yet it is typically represented as aggregated counts th

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Rethinking LoRA Memory Through the Lens of KV Cache Compression

DGX agent

arXiv:2606.05698v1 Announce Type: new Abstract: Parametric retrieval augmentation encodes document information into lightweight, document-specific modules such as LoRA adapters, reducing the need to i

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Ten Headache Specialists versus Artificial Intelligence for Clinical Literature Summarization: A Critical Evaluation and Comparison

DGX agent

arXiv:2606.05436v1 Announce Type: cross Abstract: Summarizing the latest medical literature to guide clinical decision-making is essential for evidence-based medicine and high-quality patient care. Ye

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Token costs are why there will be no saas apocalypse / good dev tools are cached intelligence for agents! The popular theory goes: agents ca…

DGX agent

Token costs are why there will be no saas apocalypse / good dev tools are cached intelligence for agents! The popular theory goes: agents can write code, so they'll just rebuild every tool from scratc

model-releasesclem-delangue--x
5 Jun 2026
Model Releases

Towards One-to-Many Temporal Grounding

DGX agent

arXiv:2606.06294v1 Announce Type: new Abstract: Temporal Grounding (TG) aims to localize video segments corresponding to a textual query. Prior research predominantly focuses on single-segment retriev

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

Unlocking dependable responses with Gemini Enterprise Agent Platform’s Agentic RAG

DGX agent

Google's RAG Engine securely connects private enterprise data to LLMs to improve answer accuracy and reduce hallucinations , making it a key component of the Gemini Enterprise Agent Platform for build

model-releasesgoogle-research
5 Jun 2026
Research

Unsupervised Monocular 3D Keypoint Discovery from Multi-View Diffusion Priors

DGX agent

arXiv:2507.12336v2 Announce Type: replace Abstract: Most existing 3D keypoint estimation methods rely on manual annotations or calibrated multi-view images, both of which are expensive to collect. Thi

researcharxiv-cs-cv
5 Jun 2026
Research

USAD 2.0: Scaling Representation Distillation for Universal Audio Understanding

DGX agent

arXiv:2606.06444v1 Announce Type: cross Abstract: Audio encoders are critical to modern audio applications as large language models (LLMs) increasingly rely on a single encoder for diverse inputs. Whi

researcharxiv-cs-cl
5 Jun 2026
Local Ai

VASO: Formally Verifiable Self-Evolving Skills for Physical AI Agents

DGX agent

arXiv:2606.05395v1 Announce Type: new Abstract: Reusable robot skills are becoming the basic units through which embodied agents turn open-ended instructions into long-horizon physical behavior. We ar

local-aiarxiv-cs-ro
5 Jun 2026
Model Releases

Video-Rate Streaming Stylization on a Vision-Aware MLLM-Conditioned Edit Diffusion: Asymmetric Batched Inference on a Distilled UNet + MLLM Text Encoder

DGX agent

arXiv:2606.05981v1 Announce Type: new Abstract: Aggressive distillation of the diffusion U-Net inverts the per-frame bottleneck of real-time text-to-image pipelines: once the denoiser is a 4-step or 1

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

VideoKR: Towards Knowledge- and Reasoning-Intensive Video Understanding

DGX agent

arXiv:2606.05259v1 Announce Type: new Abstract: We introduce VideoKR, the first large-scale training corpus specifically designed to strengthen knowledge- and reasoning-intensive video understanding.

model-releasesarxiv-cs-cv
5 Jun 2026
Local Ai

Vision Hopfield Memory Networks

DGX agent

arXiv:2603.25157v2 Announce Type: replace-cross Abstract: Recent vision and multimodal foundation backbones, such as Transformer families and state-space models like Mamba, have achieved remarkable pr

local-aiarxiv-cs-cv
5 Jun 2026
Research

Visual Commonsense Driven Knowledge Refinements for Scene Graph Generation

DGX agent

arXiv:2606.06369v1 Announce Type: new Abstract: Learning-driven Scene Graph Generation (SGG) models excel on frequent relation types but degrade sharply under annotation sparsity, failing to capture r

researcharxiv-cs-cv
5 Jun 2026
Model Releases

VTI-CoT: Visual-Textual Interleaved Chain of Thought for Video Reasoning

DGX agent

arXiv:2606.05736v1 Announce Type: new Abstract: Video reasoning aims to understand complex temporal events and causal relationships within videos. Recently, Chain-of-Thought (CoT) has been introduced

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

We've made a breakthrough in self-evolving AI scientists moving from 'search' to 'principled discovery': Scientific discovery requires that …

DGX agent

We've made a breakthrough in self-evolving AI scientists moving from 'search' to 'principled discovery': Scientific discovery requires that the search space itself changes, and an AI scientist must pe

model-releasesgary-marcus--x
5 Jun 2026
Model Releases

Your AI chatbot is only as good as the data behind it. This n8n template from our friends at @apify shows you how to wire up a RAG pipeline …

DGX agent

Your AI chatbot is only as good as the data behind it. This n8n template from our friends at @apify shows you how to wire up a RAG pipeline using Apify + Pinecone + Gemini so your chatbot can answer q

model-releasespinecone--x
5 Jun 2026
Model Releases

Affordance2Action: Task-Conditioned Scene-level Affordance Grounding for Real-Time Manipulation

DGX agent

arXiv:2606.04172v1 Announce Type: new Abstract: Task-conditioned manipulation requires grounding instructions to task-relevant functional parts rather than object categories. This setting is scene-dep

model-releasesarxiv-cs-ro
4 Jun 2026
Model Releases

Agent Planning Benchmark: A Diagnostic Framework for Planning Capabilities in LLM Agents

DGX agent

arXiv:2606.04874v1 Announce Type: new Abstract: Planning is central to LLM agents: before acting, an agent must decompose goals, select tools, reason over constraints, and decide when a task is infeas

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

ALINC: Active Learning for Inductive Node Classification via Graph Sampling

DGX agent

arXiv:2606.04647v1 Announce Type: new Abstract: Active learning (AL) for node classification typically focuses on selecting the most informative nodes for annotation within one or a few large graphs (

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Andon Labs' Real-World AI Evals: Claude calls the FBI, AI CEOs, price cartels, Butter-Bench, & Luna https://latent.space/p/andon @andonlabs …

DGX agent

Andon Labs' Real-World AI Evals: Claude calls the FBI, AI CEOs, price cartels, Butter-Bench, & Luna https://latent.space/p/andon @andonlabs cofounders @lukaspet and @axelbacklund explain why dollar-de

model-releasesswyx--x
4 Jun 2026
Model Releases

Benchmarking Living-Screen-Native GUI Agents on Short-Video Platforms

DGX agent

arXiv:2606.04701v1 Announce Type: cross Abstract: GUI agents today assume a static screen, where the world is frozen between two actions. However, real interfaces such as short-video applications viol

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Beyond Structural Symmetries: Linear Mode Connectivity via Neuron Identifiability

DGX agent

arXiv:2606.04754v1 Announce Type: new Abstract: Many striking phenomena in deep learning, such as linear mode connectivity and the structured behavior of training dynamics, are closely tied to paramet

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Breaking Bad Molecules: Are MLLMs Ready for Structure-Level Molecular Detoxification?

DGX agent

arXiv:2506.10912v4 Announce Type: replace Abstract: Toxicity remains a leading cause of early-stage drug development failure. Despite advances in molecular design and property prediction, the task of

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Can Generalist Agents Automate Data Curation?

DGX agent

arXiv:2606.04261v1 Announce Type: new Abstract: Curating training data is among the most consequential yet labor-intensive parts of modern AI development: practitioners iteratively propose, implement,

model-releasesarxiv-cs-ai
4 Jun 2026
Research

Can Reasoning Path still be Effective as Input? Bridging Post-Reasoning to Chain-of-Thought Compression

DGX agent

arXiv:2510.08647v2 Announce Type: replace-cross Abstract: Recent developments have enabled advanced reasoning in Large Language Models (LLMs) via long Chain-of-Thought (CoT), trading efficiency during

researcharxiv-cs-ai
4 Jun 2026
Model Releases

CDPM-Align: Multi-Scale Guidance-Aligned Diffusion Pretraining for Robust Few-Shot Anatomical Landmark Detection

DGX agent

arXiv:2606.04898v1 Announce Type: new Abstract: Anatomical landmark detection is a fundamental task in medical image analysis supporting a wide range of diagnostic and interventional workflows. Althou

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

ChannelTok: Efficient Flexible-Length Vision Tokenization

DGX agent

arXiv:2606.04461v1 Announce Type: new Abstract: Leading flexible vision tokenizers achieve SOTA quality at an extreme cost, relying on parameter-heavy backbones and slow, multi-step generative decoder

model-releasesarxiv-cs-cv
4 Jun 2026
Tutorials

Characterizing, Evaluating, and Optimizing Complex Reasoning

DGX agent

arXiv:2602.08498v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) increasingly rely on reasoning traces with complex internal structures. However, existing work lacks a unified answer

tutorialsarxiv-cs-cl
4 Jun 2026
Research

ChatSOP: An SOP-Guided MCTS Planning Framework for Controllable LLM Dialogue Agents

DGX agent

arXiv:2407.03884v4 Announce Type: replace-cross Abstract: Dialogue agents powered by Large Language Models (LLMs) show superior performance in various tasks. Despite the better user understanding and

researcharxiv-cs-ai
4 Jun 2026
Research

Computational conceptual history of scientific concepts: From early digital methods to LLMs

DGX agent

arXiv:2606.04118v1 Announce Type: new Abstract: This article situates large language models (LLMs) within the longer history of computational approaches to concept analysis in the history, philosophy,

researcharxiv-cs-cl
4 Jun 2026
Safety

Confidence Before Answering: A Paradigm Shift for Efficient LLM Uncertainty Estimation

DGX agent

arXiv:2603.05881v2 Announce Type: replace Abstract: Reliable deployment of large language models (LLMs) requires accurate uncertainty estimation. Existing methods are predominantly answer-first, produ

safetyarxiv-cs-cl
4 Jun 2026
Model Releases

Constraint-Enhanced Physical Search through Correlation Matching

DGX agent

arXiv:2606.03554v1 Announce Type: cross Abstract: Physical systems do not merely add noise to search processes; they impose constraints that generate structured correlations. We propose a principle of

model-releasesarxiv-cs-ai
4 Jun 2026
Research

Cross-Prompt Generalization in Detecting AI-Generated Fake News Using Interpretable Linguistic Features

DGX agent

arXiv:2606.04199v1 Announce Type: new Abstract: The increasing use of large language models has raised concerns about the spread of AI-generated fake news, particularly under varying prompting strateg

researcharxiv-cs-cl
4 Jun 2026
Model Releases

D^3-MoE:Dual Disentangled Diffusion Mixture-of-Experts for Style-Controllable End-to-End Autonomous Driving

DGX agent

arXiv:2606.04884v1 Announce Type: new Abstract: Traditional end-to-end autonomous driving frameworks frequently suffer from the 'style-averaging' dilemma when trained on high-variance human demonstrat

model-releasesarxiv-cs-ro
4 Jun 2026
Model Releases

DLLG: Dynamic Logit-Level Gating of LLM Experts

DGX agent

arXiv:2606.04378v1 Announce Type: new Abstract: Leveraging multiple specialized LLMs can combine complementary strengths, but existing approaches trade adaptability for stability: routing commits prem

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

DLO-Lab: Benchmarking Deformable Linear Object Manipulations with Differentiable Physics

DGX agent

arXiv:2606.04206v1 Announce Type: new Abstract: We address the challenge of enabling robots to manipulate deformable linear objects (DLOs), such as ropes, cables, and rubber bands. Prior work has prim

model-releasesarxiv-cs-ro
4 Jun 2026
Model Releases

Emotion Entanglement and Bayesian Inference for Multi-Dimensional Emotion Understanding

DGX agent

arXiv:2604.00819v2 Announce Type: replace-cross Abstract: Understanding emotions in natural language is inherently a multi-dimensional reasoning problem, where multiple affective signals interact thro

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

From Untrusted Input to Trusted Memory: A Systematic Study of Memory Poisoning Attacks in LLM Agents

DGX agent

arXiv:2606.04329v1 Announce Type: cross Abstract: Memory is a core component of AI agents, enabling them to accumulate knowledge across interactions and improve performance. However, persistent memory

model-releasesarxiv-cs-ai
4 Jun 2026
← Previous
1…692693694695696…1359
Next →