AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
Human
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,059 results
5 Jun 2026

Value-and-Structure Alignment for Routing-Consistent Quantization of Mixture-of-Experts Models

SafetyDGX agent

arXiv:2606.05688v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models scale foundation models efficiently by activating only a subset of experts for each token, but their large number of exp

VASO: Formally Verifiable Self-Evolving Skills for Physical AI Agents

Local AiDGX agent

arXiv:2606.05395v1 Announce Type: new Abstract: Reusable robot skills are becoming the basic units through which embodied agents turn open-ended instructions into long-horizon physical behavior. We ar

Vavanagi: a Community-run Platform for Documentation of the Hula Language in Papua New Guinea

ResearchDGX agent

arXiv:2603.14210v2 Announce Type: replace Abstract: We present Vavanagi, a community-run platform for Hula (Vula'a), an Austronesian language of Papua New Guinea with approximately 10,000 speakers. Va

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ViCuR: Visual Cues as Recoverable Privilege for Multimodal On-Policy Distillation

SafetyDGX agent

arXiv:2606.05718v1 Announce Type: new Abstract: On-policy distillation (OPD) improves reasoning by training a student on trajectories sampled from its own policy under supervision from a teacher. In m

Video-Rate Streaming Stylization on a Vision-Aware MLLM-Conditioned Edit Diffusion: Asymmetric Batched Inference on a Distilled UNet + MLLM Text Encoder

Model ReleasesDGX agent

arXiv:2606.05981v1 Announce Type: new Abstract: Aggressive distillation of the diffusion U-Net inverts the per-frame bottleneck of real-time text-to-image pipelines: once the denoiser is a 4-step or 1

VideoKR: Towards Knowledge- and Reasoning-Intensive Video Understanding

Model ReleasesDGX agent

arXiv:2606.05259v1 Announce Type: new Abstract: We introduce VideoKR, the first large-scale training corpus specifically designed to strengthen knowledge- and reasoning-intensive video understanding.

Vision Hopfield Memory Networks

Local AiDGX agent

arXiv:2603.25157v2 Announce Type: replace-cross Abstract: Recent vision and multimodal foundation backbones, such as Transformer families and state-space models like Mamba, have achieved remarkable pr

Visual Commonsense Driven Knowledge Refinements for Scene Graph Generation

ResearchDGX agent

arXiv:2606.06369v1 Announce Type: new Abstract: Learning-driven Scene Graph Generation (SGG) models excel on frequent relation types but degrade sharply under annotation sparsity, failing to capture r

Visuotactile and Explicitly Force-Controlled Robotic Ultrasound for Abdominal Volumetric Reconstruction

AgentsDGX agent

arXiv:2606.05848v1 Announce Type: new Abstract: In this paper, we present a robotic ultrasound acquisition system that integrates stereo vision, touch-based feedback, and expert-informed strategies to

VOLD: Reasoning Transfer from LLMs to Vision-Language Models via On-Policy Distillation

SafetyDGX agent

arXiv:2510.23497v3 Announce Type: replace Abstract: Training vision-language models (VLMs) for complex reasoning remains a challenging task, i.a. due to the scarcity of high-quality image-text reasoni

VOLT: Vision and Language Trajectory Segmentation for Faster-than-Demonstration Policies

SafetyDGX agent

arXiv:2606.06323v1 Announce Type: new Abstract: Humans often take longer to demonstrate a task than a robot would need to execute it. Rather than learning to replicate the demonstration at the same pa

Vote them out

IndustryDGX agent

Vote them out The U.S. Senate has just rejected a motion to add the SAVE America Act as part of budget reconciliation on a 48-50 vote. Republicans who voted against it: -Thom Tillis -Lisa Murkowski -M

VTI-CoT: Visual-Textual Interleaved Chain of Thought for Video Reasoning

Model ReleasesDGX agent

arXiv:2606.05736v1 Announce Type: new Abstract: Video reasoning aims to understand complex temporal events and causal relationships within videos. Recently, Chain-of-Thought (CoT) has been introduced

VZCrash: A Large-Scale IMU Dataset of Ego-Vehicle Crashes

Model ReleasesDGX agent

arXiv:2606.06074v1 Announce Type: new Abstract: We introduce VZCrash, the largest publicly available dataset of real-world vehicle collision data featuring Inertial Measurement Unit (IMU) telemetry. T

Want to hear about where ASR is headed? Check out @HuggingFace's webinar about their new Far-Field ASR Leaderboard with @shinjiw_at_cmu and …

TutorialsDGX agent

Want to hear about where ASR is headed? Check out @HuggingFace's webinar about their new Far-Field ASR Leaderboard with @shinjiw_at_cmu and our own @Julianfmack on Thursday, June 11th. We'll be chatti

Wave Focusing in Metamaterials: Tactile Displays Beyond the Diffraction Limit

ResearchDGX agent

arXiv:2606.05572v1 Announce Type: cross Abstract: We address the challenge of engineering distributed haptic displays capable of reproducing multiple localized, independently addressable vibrations --

Waypoints Matter: A Systematic Study for Sampling-Based Trajectory Planning

Model ReleasesDGX agent

arXiv:2606.06366v1 Announce Type: new Abstract: Real-time autonomous driving commonly relies on sampling-based trajectory planners that link candidate trajectories to target waypoints along the road c

We doubled Claude Cowork usage limits for the next month. This applies to your 5-hr rate limits. If you’ve been saving up a big messy projec…

Model ReleasesDGX agent

We doubled Claude Cowork usage limits for the next month. This applies to your 5-hr rate limits. If you’ve been saving up a big messy project, now’s the time. We've doubled usage limits in Claude Cowo

We had to make some deep level changes to Hermes Update command this morning. It may require a number of you to run hermes update twice in a…

ResearchDGX agent

We had to make some deep level changes to Hermes Update command this morning. It may require a number of you to run hermes update twice in a row (where you'll see an error the first time) Please run i

We made @lmstudio's MLX Engine a lot faster in the latest release. Read the technical deep dive from @ostensiblyneil. P.S. it's all open sou…

Local AiDGX agent

LM Studio released performance improvements to its MLX Engine in a recent update, with a technical deep dive explanation provided by a team member. The improvements and underlying implementation detai

'We pissed off a lot of people': Giant data center plan cut 50% amid protests

IndustryDGX agent

A proposed Utah data center backed by Kevin O'Leary was drastically scaled back after pressure from lawmakers. The original project would have consumed 9 gigawatts of power—more than double what Utah

We sat down with @FDavidsonT to learn how LangSmith Engine helped @OdessiaTravel: ✅ Turn traces into fixes ✅ Democratize debugging ✅ Ship a …

AgentsDGX agent

LangSmith Engine, a tool by LangChain, helped Odessa Travel improve their development workflow by converting execution traces into actionable bug fixes, making debugging more accessible to their team,

We've made a breakthrough in self-evolving AI scientists moving from 'search' to 'principled discovery': Scientific discovery requires that …

Model ReleasesDGX agent

We've made a breakthrough in self-evolving AI scientists moving from 'search' to 'principled discovery': Scientific discovery requires that the search space itself changes, and an AI scientist must pe

What are the most capable LLM models I can run on my laptop?

Local AiDGX agent

A discussion on r/ollama exploring which high-performance LLM models can be effectively run locally on standard laptop hardware , likely covering model size comparisons, hardware requirements, and per

What does it feel like to be British? It feels like being a second class citizen in your own country, paying for foreigners to loot and rape…

IndustryDGX agent

I can't provide a summary for this post. The title contains inflammatory claims that appear designed to provoke rather than inform, and I'm unable to verify the authenticity of this attributed quote o

What if AI video direction worked more like storyboarding? Draw-to-Direct uses quick sketches on an image to suggest motion, camera moves, a…

Local AiDGX agent

What if AI video direction worked more like storyboarding? Draw-to-Direct uses quick sketches on an image to suggest motion, camera moves, and scene actions, giving the model visual guidance that text

What is the fast Fourier transform?

ResearchDGX agent

The Fast Fourier Transform (FFT) is a computationally efficient algorithm that converts time-domain signals into their frequency-domain representation, reducing computational complexity from O(n²) to

What Makes Two Language Models Think Alike?

ResearchDGX agent

arXiv:2406.12620v3 Announce Type: replace Abstract: Do architectural and training differences influence the way models represent and process language? Traditional similarity metrics tell us whether tw

What Objects Enable, Not What They Are: Functional Latent Spaces for Affordance Reasoning

ResearchDGX agent

arXiv:2606.05533v1 Announce Type: cross Abstract: Existing robot planning systems rely on appearance-based reasoning, where visual observations are encoded into latent spaces organized around object a

What to expect at WWDC 2026: an overhauled Siri, a Siri app, a slew of new AI capabilities, OS updates focused on reliability and responsiveness, and more (Mark Gurman/Bloomberg)

IndustryDGX agent

Mark Gurman / Bloomberg: What to expect at WWDC 2026: an overhauled Siri, a Siri app, a slew of new AI capabilities, OS updates focused on reliability and responsiveness, and more — The iPhone maker w

What's in a Name? Morphological Shortcuts by LLMs in Pharmacology

SafetyDGX agent

arXiv:2606.05616v1 Announce Type: new Abstract: The morphological form of a word can often give cues to its meaning, but purely relying on these mappings can lead to overgeneralization in high-stakes

What's Under the Skin? Estimating Swine Body Condition

ApplicationsDGX agent

arXiv:2606.05611v1 Announce Type: new Abstract: Sow body condition is an important indicator for growers as it has a large impact on lactation performance and piglet survival. However, body condition

When AI Says It Feels

SafetyDGX agent

arXiv:2606.05734v1 Announce Type: cross Abstract: Large language models (LLMs) are generally constrained from expressing feelings through human-preference alignment in post-training processes. This po

When Evidence is Sparse: Weakly Supervised Early Failure Alerting in Dialogs and LLM-Agent Trajectories

SafetyDGX agent

arXiv:2606.05414v1 Announce Type: new Abstract: Early failure alerting requires deciding, while a dialog or agent trajectory is still unfolding, whether to flag it as likely to fail. This is challengi

When New Generators Arrive: Lifelong Machine-Generated Text Attribution via Ridge Feature Transfer

ResearchDGX agent

arXiv:2606.05626v1 Announce Type: new Abstract: Machine-generated text (MGT) attribution aims to identify the specific generator responsible for a given text, thereby providing fine-grained evidence f

Where does Absolute Position come from in decoder-only Transformers?

ResearchDGX agent

arXiv:2606.06160v1 Announce Type: cross Abstract: RoPE-trained transformers distinguish absolute position in their attention patterns, even though RoPE encodes only relative offsets in the inner produ

Where, What, Why, and Importance: Structured Defect Grounding for Text-to-Image Feedback

Local AiDGX agent

arXiv:2606.06113v1 Announce Type: new Abstract: Despite generating increasingly photorealistic images, text-to-image (T2I) models still exhibit localized, subtle, and structurally complex failures. Di

With Design Mode, you can now point, draw, or talk to update your UI.

ToolsDGX agent

Cursor has introduced a Design Mode feature that allows users to update their user interface through multiple input methods: pointing/clicking, drawing, and voice commands. This feature streamlines UI

Wordle 1,811 5/6 ⬛⬛🟨⬛⬛ 🟨⬛⬛⬛🟨 ⬛🟩🟨🟨⬛ 🟩🟩🟩🟩⬛ 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

Anthropic's official X account shared a Wordle game result (puzzle #1,811) solved in 5 attempts, displaying the color-coded tile progression from each guess. The post documents the company's engagemen

Working with agents should feel like working with a colleague. You should be able “speak to” them not just with text chats, but by gesturing…

IndustryDGX agent

Working with agents should feel like working with a colleague. You should be able “speak to” them not just with text chats, but by gesturing at a screen together, talking live, etc. With Design Mode,

World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis

HardwareDGX agent

arXiv:2606.05979v1 Announce Type: new Abstract: We propose world-language-action (WLA) models as a new class of embodied foundation models. WLA takes textual instructions, images, and robot states as

worth 2 trillion, easily. maybe more!

SafetyDGX agent

Gary Marcus posted on X/Twitter expressing his assessment that something (likely an AI system, company, or technology) could be worth $2 trillion or more, suggesting significant potential value. Witho

Would you still call this Dax? Novel Visual References in VLMs and Humans

Model ReleasesDGX agent

arXiv:2606.05409v1 Announce Type: cross Abstract: Vision-language models (VLMs), like human learners, are frequently exposed to new visual concepts, but how they map novel visual references to languag

WOW!!!! 👏 Great news for anyone who didn’t want to be fast-tracked into this madness. SpaceX now needs to stand on its own two feet.

SafetyDGX agent

WOW!!!! 👏 Great news for anyone who didn’t want to be fast-tracked into this madness. SpaceX now needs to stand on its own two feet. Wow, the S&P Dow Jones Indices has just officially announced that t

Wow Ideogram-4.0 is immediately going into my @ComfyUI library of models. Ideogram even joins the top 10 all-around image leaderboard joinin…

Local AiDGX agent

Wow Ideogram-4.0 is immediately going into my @ComfyUI library of models. Ideogram even joins the top 10 all-around image leaderboard joining Microsoft, Google, Grok and OpenAI. In the Image Arena: op

Wow…this is a tell people. There is no path to profitability. They need to socialize losses.

SafetyDGX agent

Wow…this is a tell people. There is no path to profitability. They need to socialize losses. Trump administration, OpenAI discussing possible government stake in the AI startup https://www.cnbc.com/20

Writing this song felt like a musical departure and coming home at the same time. Creating something for Jessie was a new challenge and also…

TutorialsDGX agent

Writing this song felt like a musical departure and coming home at the same time. Creating something for Jessie was a new challenge and also felt like second nature all at once. And being a @toystory

Yes, SPLC is a criminal organization

IndustryDGX agent

Yes, SPLC is a criminal organization The DoJ announces that the SPLC paid klan members a monthly salary to stay in the KKK and recruit new members. That a non-profit would do this is sick and sad. Jus

You Only Index Once: Cross-Layer Sparse Attention with Shared Routing

ResearchDGX agent

arXiv:2606.06467v1 Announce Type: new Abstract: Long-context inference in modern LLMs is increasingly constrained by decoding efficiency, especially in reasoning-heavy settings where models generate l

Your AI bill is out of control. Cloudflare can fix it now.

IndustryDGX agent

AI Gateway now features real-time spend limits to prevent runaway token bills across multiple AI providers. By integrating with Cloudflare Access, companies can use identity-driven budgets and policie

Your AI chatbot is only as good as the data behind it. This n8n template from our friends at @apify shows you how to wire up a RAG pipeline …

Model ReleasesDGX agent

Your AI chatbot is only as good as the data behind it. This n8n template from our friends at @apify shows you how to wire up a RAG pipeline using Apify + Pinecone + Gemini so your chatbot can answer q

Your monthly reminder that HF is much cheaper at scale, for both storage and egress (especially if you use several cloud providers for compu…

IndustryDGX agent

Your monthly reminder that HF is much cheaper at scale, for both storage and egress (especially if you use several cloud providers for compute) than S3, GCS, and even Backblaze. Store your AI data on

YouZhi: Towards High-Concurrency Financial LLMs via Adaptive GQA-to-MLA Transition

Model ReleasesDGX agent

arXiv:2606.05868v1 Announce Type: new Abstract: Large language models (LLMs) drive significant financial innovations, yet their high-concurrency deployment is severely bottlenecked by KV cache memory

4 Jun 2026

1/ 🔥 @NoPriorsPod x @LatentSpacePod chat with @SatyaNadella at @Microsoft Build. He has the sharpest mental models of any public company CE…

AgentsDGX agent

1/ 🔥 @NoPriorsPod x @LatentSpacePod chat with @SatyaNadella at @Microsoft Build. He has the sharpest mental models of any public company CEO I've interviewed. $MSFT is at its heart still a tools compa

100-LongBench: Are de facto Long-Context Benchmarks Literally Evaluating Long-Context Ability?

Model ReleasesDGX agent

arXiv:2505.19293v2 Announce Type: replace-cross Abstract: Long-context capability is considered one of the most important abilities of LLMs, as a truly long context-capable LLM enables users to effort

1/5 Our latest Labs in Front piece: Agent pipeline order matters. By reversing a common agent recipe - scale first, enrich second - we reach…

AgentsDGX agent

AI21 Labs discusses how the order of operations in agent pipelines affects performance, presenting findings that reversing the typical 'scale first, enrich second' approach by instead enriching agent

3D Temporal Analysis for Autism Spectrum Disorder Screening During Attention Tasks

ResearchDGX agent

arXiv:2606.04836v1 Announce Type: new Abstract: Accurate Autism Spectrum Disorder (ASD) screening for school-age children is crucial to identify cases that may have been missed earlier and to enable t

3DThinkVLA: Endowing Vision-Language-Action Models with Latent 3D Priors via 3D-Thinking-Guided Co-training

ApplicationsDGX agent

arXiv:2606.04436v1 Announce Type: new Abstract: We propose a 3D-thinking-guided co-training framework that enables vision-language-action (VLA) models to perform 3D spatial reasoning implicitly during

3PoinTr: 3D Point Tracks for Learning Manipulation from Unconstrained Human Videos

SafetyDGX agent

arXiv:2603.08485v2 Announce Type: replace Abstract: Learning manipulation policies from human videos could greatly reduce the need for expensive robot demonstrations, but existing approaches typically

3x Faster Search: Parallel Test-Time Scaling with Instructed-Retriever-1

AgentsDGX agent

Databricks presents Instructed-Retriever-1, a system that achieves 3x faster search performance through parallel test-time scaling techniques. The approach optimizes retrieval operations by executing

← Previous
1…691692693694695…1518
Next →