AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,060 results
5 Jun 2026

IR3DE: A Linear Router for Large Language Models

ResearchDGX agent

arXiv:2606.06098v1 Announce Type: new Abstract: Foundational Large Language Models (LLMs) demonstrate proficiency on a wide range of general tasks, and achieve remarkable results on various specialize

Is Diversity All You Need for Scalable Robotic Manipulation?

SafetyDGX agent

arXiv:2507.06219v2 Announce Type: replace Abstract: Data scaling has driven remarkable success in foundation models for Natural Language Processing (NLP) and Computer Vision (CV), yet the principles o

Is This Edit Correct? A Multi-Dimensional Benchmark for Reasoning-Aware Image Editing

Model ReleasesDGX agent

arXiv:2606.05172v1 Announce Type: cross Abstract: Diffusion-based image editing has achieved strong visual fidelity under natural language instructions, yet most existing systems still operate at the

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Isomorphic likely that every single person in that game was already familiar with this

ToolsDGX agent

Isomorphic likely that every single person in that game was already familiar with this Levels of Mafia/Werewolf/Secret Hitler: • You play, poorly • You play, skillfully • You play, poorly, so that peo

It was a pleasure to discuss all of these topics on @NxThompson's podcast, The Most Interesting Thing in AI. Thanks for having me on! https:…

SafetyDGX agent

Yoshua Bengio appeared as a guest on 'The Most Interesting Thing in AI' podcast hosted by N.X. Thompson, where he discussed various topics related to artificial intelligence. The appearance was shared

It's the time of the year when you shouldn't come to New York if you are not planning on moving permanently.

IndustryDGX agent

It's the time of the year when you shouldn't come to New York if you are not planning on moving permanently. It's the time of the year when you shouldn't come to New York if you are not planning on mo

Join us now!

Local AiDGX agent

This is a social media post from ComfyUI's X (Twitter) account inviting users to join their community or platform. Without access to the specific post content, it likely promotes ComfyUI's node-based

Join us on a live interview with the CEO of ComfyUI!

Local AiDGX agent

Join us on a live interview with the CEO of ComfyUI! Today on TWiST, we're joined by @yoland_yan, Founder/ CEO of @ComfyUI. With 4M users, 150K downloads a day, and a $500M valuation from Craft, @Comf

⚠️ Keep your eye on the ball, and don’t panic over Anthropic’s new blog. Here’s why: Anthropic is trying to strike terror into everyone’s he…

SafetyDGX agent

⚠️ Keep your eye on the ball, and don’t panic over Anthropic’s new blog. Here’s why: Anthropic is trying to strike terror into everyone’s hearts – “full recursive self-improvement also might increase

Knowledge Distillation for Visual Autoregressive Models

ResearchDGX agent

arXiv:2606.06078v1 Announce Type: new Abstract: Autoregressive (AR) image generation models are highly expressive but computationally intensive, motivating effective model compression. Knowledge disti

KV-Control: Parameter-Efficient K/V Injection for Trajectory-Controlled Text-to-Motion

Model ReleasesDGX agent

arXiv:2606.05624v1 Announce Type: new Abstract: Text-conditioned 3D human motion models now synthesize plausible motions from prompts, but practical animation and embodied-agent workflows rarely stop

L-SDPPO: Policy Optimization of Spiking Diffusion Policy for Intra-vehicular Robotic Manipulation

SafetyDGX agent

arXiv:2606.06049v1 Announce Type: new Abstract: Intra-vehicular robots in spacecraft help reduce astronaut workload and improve mission efficiency. Recent research focuses on using deep learning metho

LadderMan: Learning Humanoid Perceptive Ladder Climbing

SafetyDGX agent

arXiv:2606.05873v1 Announce Type: cross Abstract: Humanoid robots hold great promise for operating in human-centered environments, yet ladder climbing remains one of the most challenging tasks due to

LANTERN: Layered Archival and Temporal Episodic Retrieval Network for Long-Context LLM Conversations

ApplicationsDGX agent

arXiv:2606.05182v1 Announce Type: new Abstract: Large language models discard critical details when conversation history is compacted to fit within finite context windows. We present LANTERN (Layered

Large Language Models are Perplexed by some Political Parties

SafetyDGX agent

arXiv:2606.05937v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used, including in political applications, but their political fairness has been little studied. We assess

Last night, I read the entirety of C.S. Lewis' The Screwtape Letters. It's a novel told in the form of letters written by a demon to another…

IndustryDGX agent

Last night, I read the entirety of C.S. Lewis' The Screwtape Letters. It's a novel told in the form of letters written by a demon to another demon instructing him on ways to manipulate his 'patient' t

Last year: people had to steal GPU’s from armored trucks. This year: SpaceX is leasing to GPUs left, right, and center, because Xai couldn’t…

HardwareDGX agent

Last year: people had to steal GPU’s from armored trucks. This year: SpaceX is leasing to GPUs left, right, and center, because Xai couldn’t figure out what to do with them. SpaceX just disclosed a ne

Latent Implicit Visual Reasoning

ResearchDGX agent

arXiv:2512.21218v2 Announce Type: replace Abstract: While Large Multimodal Models (LMMs) have made significant progress, they remain largely text-centric, relying on language as their core reasoning m

Latent Reasoning with Normalizing Flows

SafetyDGX agent

arXiv:2606.06447v1 Announce Type: new Abstract: Large language models often improve reasoning by generating explicit chain-of-thought (CoT), demonstrating the importance of intermediate computation. H

LatentSkill: From In-Context Textual Skills to In-Weight Latent Skills for LLM Agents

Model ReleasesDGX agent

arXiv:2606.06087v1 Announce Type: new Abstract: Agent systems increasingly use textual skills to encode reusable task procedures, but injecting these skills into the prompt at every step incurs substa

LeanMarathon: Toward Reliable AI Co-Mathematicians through Long-Horizon Lean Autoformalization

Local AiDGX agent

arXiv:2606.05400v1 Announce Type: cross Abstract: Long-horizon autoformalization of research mathematics fails not only at hard lemmas, but at scale: statements drift, dependencies tangle, context dec

Learning Contact Representation for Leg Odometry

ResearchDGX agent

arXiv:2606.05501v1 Announce Type: new Abstract: The estimation of odometry in legged robots depends on the assumption that the velocity of the foot with respect to the world remains zero during the st

Learning from Demonstrations over Riemannian Manifolds using Neural ODEs: An Extended Abstract

ResearchDGX agent

arXiv:2606.05422v1 Announce Type: new Abstract: Learning from demonstratins (LfD) is usually performed over Euclidean spaces, while the robot state, e.g. orientation, naturally evolves over curved spa

Learning Geometric Representations from Videos for Spatial Intelligent Multimodal Large Language Models

ApplicationsDGX agent

arXiv:2606.05833v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) excel at 2D semantic understanding but lack intrinsic 3D awareness, resulting in representations that fail to m

Learning of Robot Safety Policies via Adversarial Synthetic Scenarios

SafetyDGX agent

arXiv:2606.05952v1 Announce Type: new Abstract: In this work, we propose an agentic gamification framework for hazard-informed learning of robot safety policies through synthetic scenarios. We model s

Learning Predictive Visuomotor Coordination

ApplicationsDGX agent

arXiv:2503.23300v2 Announce Type: replace Abstract: Understanding and predicting human visuomotor coordination is crucial for applications in robotics, human-computer interaction, and assistive techno

Learning Self-Correction in Vision-Language Models via Rollout Augmentation

TutorialsDGX agent

arXiv:2602.08503v2 Announce Type: replace-cross Abstract: Self-correction is essential for solving complex reasoning problems in vision-language models (VLMs). However, existing reinforcement learning

Learning to Route LLMs from Implicit Cost-Performance Preferences via Meta-Learning

ResearchDGX agent

arXiv:2606.06178v1 Announce Type: cross Abstract: Large language models (LLMs) present a trade-off between performance and cost, where more powerful models incur greater expense. LLM routing aims to m

Learning Visual Spatial Planning from Symbolic State via Modality-Gap-Aware Self-Distillation

SafetyDGX agent

arXiv:2606.06076v1 Announce Type: cross Abstract: While vision-language models excel at general multimodal understanding, they still struggle with visual spatial planning. We attribute this to a perce

Learning What to Forget: Improving LLM Unlearning via Learned Token-Level Importance

ResearchDGX agent

arXiv:2606.06320v1 Announce Type: cross Abstract: Machine unlearning aims to remove targeted knowledge from a trained model while preserving its general capabilities. For autoregressive language model

Less is MoE: Trimming Experts in Domain-Specialist Language Models

Model ReleasesDGX agent

arXiv:2606.05538v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models achieve strong performance through conditional computation, but their large parameter footprint poses deployment chall

Let It Be Simple: One-Step Action Generation for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.05737v1 Announce Type: new Abstract: Diffusion-based vision-language-action (VLA) models often inherit the image-generation view: actions are generated by iterative denoising. We argue that

Let’s go 🚀

ResearchDGX agent

Let’s go 🚀 Building AI that Builds AI: Introducing the Sakana AI RSI Lab 🚀 https://sakana.ai/rsi-lab Today, we are announcing the Sakana AI Recursive Self-Improvement (RSI) Lab: a dedicated research g

Leveraging Large Language Models for Generating Research Topic Ontologies: A Multi-Disciplinary Study

ResearchDGX agent

arXiv:2508.20693v2 Announce Type: replace-cross Abstract: Ontologies and taxonomies of research fields are critical for managing and organising scientific knowledge, as they facilitate efficient class

LiAuto-GeoX: Efficient Grounded Driving Transformer

Model ReleasesDGX agent

arXiv:2606.05774v1 Announce Type: new Abstract: Dense 3D reconstruction has demonstrated immense potential for spatial understanding, yet its viability as a real-time, onboard representation for auton

Lightricks to split into two companies as it cuts another 75 jobs

Local AiDGX agent

Lightricks is laying off 75 employees (17% of workforce) as it continues restructuring amid AI disruption, less than six months after a previous round of 85 layoffs. The company is splitting into two

LightVesselNet: An Ultra-Lightweight Sub-100K Parameter Network for Retinal Blood Vessel Segmentation

Model ReleasesDGX agent

arXiv:2606.05354v1 Announce Type: new Abstract: Retinal blood vessel segmentation plays a vital role in the early detection of diabetic retinopathy and glaucoma. While recent deep learning models have

Literally the only reason to bailout OpenAI.

SafetyDGX agent

Gary Marcus argues for a specific rationale supporting a potential OpenAI bailout, though the exact reasoning is not detailed in the available information. Given Marcus's background as an AI researche

LLM-Conditioned Synthesis of Pathological Gaits via Structured Gait-Language Representations

ResearchDGX agent

arXiv:2606.06048v1 Announce Type: new Abstract: Pathological gait datasets remain scarce due to privacy, recruitment, cost, and movement variability. Our work presents a multimodal LLM-guided framewor

LLM-Enhanced Dialogue Management for Full-Duplex Spoken Dialogue Systems

ResearchDGX agent

arXiv:2502.14145v3 Announce Type: replace Abstract: Achieving full-duplex communication in spoken dialogue systems (SDS) requires real-time coordination between listening, speaking, and thinking. This

LLM-Guided ANN Index Optimization for Human-Object Interaction Retrieval

Model ReleasesDGX agent

arXiv:2606.05489v1 Announce Type: new Abstract: Retrieval systems underpin modern AI applications -- spanning visual search, recommendation engines, and multi-modal question answering. Modern multi-st

LLMs Can Leak Training Data But Do They Want To? A Propensity-Aware Evaluation of Memorization in LLMs

ResearchDGX agent

arXiv:2606.06286v1 Announce Type: new Abstract: Large language models can reproduce training data, but existing memorization evaluations mostly measure whether models can be forced to do so, rather th

LLMs do not teach people to think critically. Colleges do.

ApplicationsDGX agent

LLMs do not teach people to think critically. Colleges do. Hot take: Universities charge $300,000 for a degree that teaches you skills any LLM can do for free. At some point we need to have an honest

Localizing Prompt Ambiguity in Large Language Models with Probe-Targeted Attribution

Model ReleasesDGX agent

arXiv:2606.05486v1 Announce Type: new Abstract: Prompt ambiguity is a common source of failure in large language models, but is difficult to localize because it is a latent property of the prompt, whi

LongSpace: Exploring Long-Horizon Spatial Memory from Perception to Recall in Video

Model ReleasesDGX agent

arXiv:2606.05677v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have advanced image and video understanding and can increasingly handle longer visual inputs. Long-horizon ta

LoomVideo: Unifying Multimodal Inputs into Video Generation and Editing

Model ReleasesDGX agent

arXiv:2606.06042v1 Announce Type: new Abstract: Developing unified video generation and editing models capable of interpreting interleaved multimodal inputs is a promising yet challenging frontier fie

LoRi: Low-Rank Distillation for Implicit Reasoning

Model ReleasesDGX agent

arXiv:2606.05315v1 Announce Type: new Abstract: Implicit chain-of-thought (iCoT) methods aim to internalize reasoning in large language models, but often underperform explicit CoT prompting. We empiri

Love this. Juan Hernandez immigrates to the US, learns how to weld because it pays well, begins contracting at SpaceX, gets full time role +…

TutorialsDGX agent

Love this. Juan Hernandez immigrates to the US, learns how to weld because it pays well, begins contracting at SpaceX, gets full time role + $10,000 equity. It’s now worth just under a million dollars

Luca (@agi2asi) built Grid to centralize 12 different Google Drives into one hub. With Replit, he turned a prompt into an AI-powered employe…

ToolsDGX agent

Luca (@agi2asi) built Grid to centralize 12 different Google Drives into one hub. With Replit, he turned a prompt into an AI-powered employee hub in just 10 minutes. Now, team members can ask about an

Luca built Grid to centralize 12 different Google Drives into one hub. With Replit, he turned a prompt into an AI-powered employee hub in ju…

ToolsDGX agent

Luca built Grid to centralize 12 different Google Drives into one hub. With Replit, he turned a prompt into an AI-powered employee hub in just 10 minutes. Now, team members can ask about anything from

Many Circuits, One Mechanism: Input Variation and Evaluation Granularity in Circuit Discovery

ResearchDGX agent

arXiv:2606.06267v1 Announce Type: new Abstract: Circuit discovery methods identify subgraphs that explain specific model behaviors, and structural differences between discovered circuits are commonly

MARDoc: A Memory-Aware Refinement Agent Framework for Multimodal Long Document QA

AgentsDGX agent

arXiv:2606.05749v1 Announce Type: new Abstract: Iterative retrieval-reasoning agents have recently shown promise for multimodal long-document question answering. However, most existing systems maintai

Marvell and Flex, a contract manufacturer for electronics, will join the S&P 500; MRVL jumps 6%+ after hours after closing down 16.74% amid a broader sell-off (Kif Leswing/CNBC)

IndustryDGX agent

Kif Leswing / CNBC: Marvell and Flex, a contract manufacturer for electronics, will join the S&P 500; MRVL jumps 6%+ after hours after closing down 16.74% amid a broader sell-off — - Marvell Technolog

MASF: A Multi-Model Adaptive Selection Framework for Abstractive Text summarization

ResearchDGX agent

arXiv:2606.05494v1 Announce Type: new Abstract: Automatic text summarization has become increasingly important due to the rapid growth of digital textual information. This paper presents a Multi-Model

Massive output uptick due to agentic AI. Complete flat adoption.

AgentsDGX agent

Jeremy Howard discusses a significant increase in output capabilities driven by agentic AI systems, while noting that adoption rates remain completely flat across the board. The observation suggests a

MAviS: A Multimodal Conversational Assistant For Avian Species

Model ReleasesDGX agent

arXiv:2603.07294v2 Announce Type: replace Abstract: Fine-grained understanding and species-specific multimodal question answering are vital for advancing biodiversity conservation and ecological monit

MCBench: A Multicontext Safety Assessment Benchmark for Omni Large Language Models

Model ReleasesDGX agent

arXiv:2606.05177v1 Announce Type: new Abstract: Existing multimodal safety benchmarks focus solely on visual inputs and cannot assess Omni Large Language Models (LLMs) that process vision, audio, and

MDP-GRPO: Stabilized Group Relative Policy Optimization for Multi-Constraint Instruction Following

Model ReleasesDGX agent

arXiv:2606.06058v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards is ideal for multi-constraint instruction following, yet standard group-relative policy optimization (G

Measuring the sensitivity of LLM-based structured extraction to prompt, model, and schema choices in clinical discharge summaries

ResearchDGX agent

arXiv:2606.05970v1 Announce Type: new Abstract: Large language models are increasingly used for structured extraction from clinical free-text notes, but the sensitivity of their output to upstream con

Mechanistic Insights into Functional Sparsity in Multimodal LLMs via CoRe Heads

Local AiDGX agent

arXiv:2606.05843v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) demonstrate remarkable proficiency on complex vision-language tasks, the mechanisms by which they extract

← Previous
1…686687688689690…1518
Next →