AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,059 results
Safety

VCap: Hypergeometric Rewards for Weak-to-Strong Visual Captioning

DGX agent

arXiv:2605.28023v1 Announce Type: cross Abstract: Visual captioning requires models to capture visual content faithfully while minimizing both omission and hallucination. As the dominant paradigm for

safetyarxiv-cs-ai
28 May 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

VEOcc: Voxel-Centric Online Semantic Occupancy Prediction For Embodied Scene Understanding

DGX agent

arXiv:2605.25059v2 Announce Type: replace Abstract: Crucial for autonomous exploration, online 3D occupancy prediction and mapping incrementally constructs dense spatial representations on the fly. Ho

agentsarxiv-cs-cv
28 May 2026
Model Releases

Verifiable Benchmarking of Long-Horizon Spatial Biology

DGX agent

arXiv:2605.28065v1 Announce Type: new Abstract: AI agents are increasingly useful for biological data analysis, but existing benchmarks mostly test broad biological knowledge, executable workflows, or

model-releasesarxiv-cs-ai
28 May 2026
Safety

Verified Misguidance: Measuring Structural Citation Failures in Search-Augmented LLMs

DGX agent

arXiv:2605.28565v1 Announce Type: cross Abstract: Users of search-augmented LLMs rely on citations as evidence that responses are grounded in real sources, and rarely verify the cited pages themselves

safetyarxiv-cs-ai
28 May 2026
Model Releases

VeriTrip: A Verifiable Benchmark for Travel Planning Agents over Unstructured Web Corpora

DGX agent

arXiv:2605.28683v1 Announce Type: new Abstract: Existing benchmarks have laid the foundation for travel planning agents by establishing API-centric paradigms. However, as the capabilities of Autonomou

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

VibeSearchBench: Benchmarking Long-horizon Proactive Search in the Wild

DGX agent

arXiv:2605.27882v1 Announce Type: cross Abstract: LLM-based agents score well on search benchmarks, yet real users consistently find results unsatisfying, revealing a persistent evaluation-experience

model-releasesarxiv-cs-ai
28 May 2026
Research

ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention

DGX agent

arXiv:2602.07574v2 Announce Type: replace-cross Abstract: Modern multimodal large language models (MLLMs) adopt a unified self-attention design that processes visual and textual tokens at every Transf

researcharxiv-cs-cl
28 May 2026
Model Releases

VideoCanvas: Unified Video Completion from Arbitrary Spatiotemporal Patches via In-Context Conditioning

DGX agent

arXiv:2510.08555v2 Announce Type: replace Abstract: Existing controllable video generation methods are typically designed for rigid, task-specific settings, such as first-frame image-to-video, inpaint

model-releasesarxiv-cs-cv
28 May 2026
Research

VidPrism: Heterogeneous Mixture of Experts for Image-to-Video Transfer

DGX agent

arXiv:2605.28229v1 Announce Type: cross Abstract: With the rapid development of pre-training technologies, adapting large-scale Vision-Language Models (VLMs) for video understanding ie image-to-video

researcharxiv-cs-ai
28 May 2026
Safety

Visualizing Latent Phase Structures in Locomotion Policies: A Multi-Environment Study with Temporal Feature Extension

DGX agent

arXiv:2605.28186v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) has been shown to achieve high performance on locomotion control tasks in MuJoCo benchmarks such as HalfCheetah, Ant

safetyarxiv-cs-ai
28 May 2026
Model Releases

VITAL: Visual-Semantic Dual Supervision for Enhanced and Interpretable Latent Reasoning in Medical MLLMs

DGX agent

arXiv:2605.28422v1 Announce Type: cross Abstract: Latent reasoning enables reasoning over continuous hidden states rather than explicit tokens, avoiding the language bottleneck and inference overhead

model-releasesarxiv-cs-ai
28 May 2026
Safety

VLA-Hijack: A Transferable Patch Attack against Vision-Language-Action Models via Visual Proprioception Hijacking

DGX agent

arXiv:2605.28083v1 Announce Type: new Abstract: While Vision-Language-Action (VLA) models have emerged as powerful generalist policies, their severe vulnerability to adversarial patches significantly

safetyarxiv-cs-cv
28 May 2026
Safety

VLM-Based Advanced Rider Assistance System for Motorcycle Safety

DGX agent

arXiv:2605.27948v1 Announce Type: new Abstract: Motorcycles face disproportionately high crash risks compared to cars due to limited protection and heightened sensitivity to surface hazards, yet Advan

safetyarxiv-cs-ro
28 May 2026
Safety

VLMs May Not Globally Enhance Human Alignment over LLMs During Natural Reading

DGX agent

arXiv:2605.28818v1 Announce Type: new Abstract: Large language models (LLMs) have become increasingly useful computational models of human language processing, but it remains unclear whether vision-la

safetyarxiv-cs-cl
28 May 2026
Safety

Voluntary Collusion with Secret Tools in Competing LLM Agents

DGX agent

arXiv:2605.27593v1 Announce Type: new Abstract: Even when a tool is explicitly described as unfair and harmful to others, ostensibly safety-aligned LLM agents still voluntarily engage in secret collus

safetyarxiv-cs-ai
28 May 2026
Model Releases

VULPO: Context-Aware Vulnerability Detection via On-Policy LLM Optimization

DGX agent

arXiv:2511.11896v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently shown strong potential in vulnerability detection (VD). However, accurately detecting vulnerabiliti

model-releasesarxiv-cs-ai
28 May 2026
Industry

🟥 W toku procesu wycieka coraz więcej szczegółów. Zabójca, Vickrum Digwa, zadzwonił na policje, a nie na pogotowie i skłamał, że to Henry g…

DGX agent

🟥 W toku procesu wycieka coraz więcej szczegółów. Zabójca, Vickrum Digwa, zadzwonił na policje, a nie na pogotowie i skłamał, że to Henry go zaatakował, był pijany, obraził go rasistowsko i strącił mu

industryelon-musk--x
28 May 2026
Model Releases

wait… if most people think 5.5 is better than 4.7, i assume that’s due to terminal coding benchmark… 4.8 is still outperformed by 5.5

DGX agent

wait… if most people think 5.5 is better than 4.7, i assume that’s due to terminal coding benchmark… 4.8 is still outperformed by 5.5 Introducing Claude Opus 4.8: it builds on Opus 4.7 with sharper ju

model-releasesjeremy-howard--x
28 May 2026
Model Releases

We also shipped dynamic workflows in Claude Code (research preview), for tasks too big for one pass. Make sure to default to auto mode so Cl…

DGX agent

We also shipped dynamic workflows in Claude Code (research preview), for tasks too big for one pass. Make sure to default to auto mode so Claude isn't stopping for permissions. It's token-intensive, s

model-releasesboris-cherny--x
28 May 2026
Industry

We are starting to be quite bullish about getting in the data infrastructure business. I just cloned 68 TB (while I only have a 4TB local di…

DGX agent

We are starting to be quite bullish about getting in the data infrastructure business. I just cloned 68 TB (while I only have a 4TB local disk) to my @huggingface training bucket in 1 minute 55 second

industryclem-delangue--x
28 May 2026
Tools

We have yet to build a truly personal computer.

DGX agent

Linus Lee argues that despite decades of development, computers remain fundamentally impersonal tools designed around generic workflows rather than individual user needs and preferences. The post like

toolslinus-lee--x
28 May 2026
Model Releases

Weak Convergence Analysis of Online Neural Actor-Critic Algorithms

DGX agent

arXiv:2403.16825v2 Announce Type: replace Abstract: We prove that a single-layer neural network trained with the online actor critic algorithm converges in distribution to a random ordinary differenti

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

WeatherCity: Urban Scene Reconstruction with Controllable Multi-Weather Transformation

DGX agent

arXiv:2602.22096v2 Announce Type: replace Abstract: Editable high-fidelity 4D scenes are crucial for autonomous driving, as they can be applied to end-to-end training and closed-loop simulation. Howev

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

We're releasing Paris 2.0, which, to our knowledge, is the world's first decentralized trained video generation model. We benchmarked it aga…

DGX agent

We're releasing Paris 2.0, which, to our knowledge, is the world's first decentralized trained video generation model. We benchmarked it against a monolithic model trained on the same data and compute

model-releasesclem-delangue--x
28 May 2026
Industry

We're selectively releasing the Paris 2.0 weights and partnering with researchers and teams interested in diffusion-based video models, worl…

DGX agent

We're selectively releasing the Paris 2.0 weights and partnering with researchers and teams interested in diffusion-based video models, world models, and embodied agents. The model is on Hugging Face:

industryclem-delangue--x
28 May 2026
Applications

We're taking on the hardest problems in the real world 🏗️🚚 🛫⚛️ Today at The AI Now Summit, held at the Louvre, we announced AI solutions …

DGX agent

We're taking on the hardest problems in the real world 🏗️🚚 🛫⚛️ Today at The AI Now Summit, held at the Louvre, we announced AI solutions for aerospace, automotive, energy, and physics. Deployed in pro

applicationsmistral-ai--x
28 May 2026
Model Releases

We've raised 65 billion in Series H funding at a 965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and @sequo…

DGX agent

We've raised 65 billion in Series H funding at a 965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and @sequoia. This investment will help us advance our research and expa

model-releasessonya-huang--x
28 May 2026
Safety

What Are We Measuring in NLG? A Meta-Analysis of Evaluation Trends 2020-2025

DGX agent

arXiv:2601.07648v2 Announce Type: replace Abstract: As Natural Language Generation (NLG) dominates modern NLP, scalable evaluation remains a critical bottleneck. Consequently, LLM-as-a-judge (LaaJ) ad

safetyarxiv-cs-cl
28 May 2026
Safety

What Frozen VLAs Already Know About Success: A Probing Study of Value-Like Structure in Foundation Robot Policies

DGX agent

arXiv:2605.28527v1 Announce Type: new Abstract: Vision--language--action (VLA) policies are trained to imitate actions; their loss never asks them to estimate reward, progress, or future success. Thei

safetyarxiv-cs-ro
28 May 2026
Model Releases

What-If World: A Causal Benchmark for General World Models in Embodied Scenarios

DGX agent

arXiv:2605.27589v1 Announce Type: new Abstract: Video generation models are increasingly used as world simulators for tasks like driving and robotic manipulation. What matters in these settings is not

model-releasesarxiv-cs-cv
28 May 2026
Applications

What to expect during Snowflake Summit: Join theCUBE June 2-3

DGX agent

The importance of Snowflake Inc. in the enterprise AI ecosystem is not necessarily its contribution to the analytics warehouse. What will be more significant going forward is Snowflake’s role as an en

applicationssiliconangle
28 May 2026
Research

When Confidence Misleads: Suffix Anchoring and Anchor-Proximity Confidence Modulation for Diffusion Language Models

DGX agent

arXiv:2605.28181v1 Announce Type: new Abstract: Diffusion language models decode text by iteratively denoising masked token sequences, making the choice of which positions to decode a central inferenc

researcharxiv-cs-cl
28 May 2026
Model Releases

When Context Flips, Safety Breaks: Diagnosing Brittle Safety in Aligned Language Models

DGX agent

arXiv:2605.27851v1 Announce Type: new Abstract: Safety benchmark scores provide incomplete evidence of deployment readiness: aligned language models often adhere to rigid rules even when a situational

model-releasesarxiv-cs-ai
28 May 2026
Research

When Discourse Pressures Conflict: Information Structure in Vision-Language Model Outputs

DGX agent

arXiv:2605.28346v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly evaluated for whether they identify the right visual content, but little is known about whether they expr

researcharxiv-cs-cl
28 May 2026
Model Releases

When do complex-valued neural networks help? A study of representation, geometry, and optimization

DGX agent

arXiv:2605.27673v1 Announce Type: new Abstract: Complex-valued Neural Networks (CVNNs) are often motivated by domains where information is naturally encoded in magnitude and phase. Yet complex-valued

model-releasesarxiv-cs-lg
28 May 2026
Agents

When Does Memory Help Multi-Trajectory Inference for Tool-Use LLM Agents?

DGX agent

arXiv:2605.28224v1 Announce Type: new Abstract: Multi-trajectory inference for tool-use LLM agents - generating multiple reasoning attempts and selecting among them - benefits from transferring knowle

agentsarxiv-cs-ai
28 May 2026
Research

When Helpful Context Leaks: Privacy Risks in Domain-Adapted ASR

DGX agent

arXiv:2605.28211v1 Announce Type: new Abstract: SpeechLLMs are increasingly deployed in professional settings where domain customisation is standard practice: users supply context in prompts with sens

researcharxiv-cs-cl
28 May 2026
Model Releases

When Interpretability Is Unequally Distributed: Fairness in Hybrid Interpretable Models

DGX agent

arXiv:2605.28626v1 Announce Type: new Abstract: Hybrid interpretable models combine a transparent component with a black-box model by assigning some examples to the former and deferring the rest to th

model-releasesarxiv-cs-lg
28 May 2026
Local Ai

When NPUs Are Not Always Faster: A Stage-Level Analysis of Mobile LLM Inference

DGX agent

arXiv:2605.27435v1 Announce Type: cross Abstract: Deploying large language models (LLMs) on mobile devices increasingly relies on heterogeneous execution, yet no prior study has systematically charact

local-aiarxiv-cs-ai
28 May 2026
Safety

When pre-training hurts LoRA fine-tuning: a dynamical analysis via single-index models

DGX agent

arXiv:2602.02855v2 Announce Type: replace Abstract: Pre-training on a source task is usually expected to facilitate fine-tuning on similar downstream problems. In this work, we mathematically show tha

safetyarxiv-cs-lg
28 May 2026
Research

When prompt perturbations break your A/B test: A valid statistical test for generative surveying

DGX agent

arXiv:2605.27463v1 Announce Type: cross Abstract: Generative surveying -- where collections of LLM-based personas provide feedback on messages -- has emerged as a cheap and scalable alternative to tra

researcharxiv-cs-ai
28 May 2026
Research

When Seekers Are Hard to Help: Evaluating Emotional Support Dialogue Systems in Worst-Case Interactions

DGX agent

arXiv:2605.28228v1 Announce Type: new Abstract: Emotional Support Dialogue Systems (ESDSes) are increasingly evaluated and trained with LLM-simulated seekers. However, such simulated seekers often beh

researcharxiv-cs-cl
28 May 2026
Safety

When Think-with-Image Meets Safety: What Determines Multimodal Jailbreak Robustness?

DGX agent

arXiv:2605.27932v1 Announce Type: cross Abstract: Think-with-image reasoning is emerging as a new inference paradigm for large vision-language models, but its safety implications remain poorly underst

safetyarxiv-cs-ai
28 May 2026
Local Ai

Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models

DGX agent

arXiv:2605.27997v1 Announce Type: cross Abstract: Large language models frequently generate toxic, hateful, or harmful content, yet existing mitigation methods rely on costly retraining or output-leve

local-aiarxiv-cs-ai
28 May 2026
Research

Where LLM Annotators Fail: Label-Free Learning on Graphs with LLMs

DGX agent

arXiv:2605.27913v1 Announce Type: new Abstract: Node classification on graphs often requires labeled nodes, yet obtaining labels at graph scale is expensive. When node attributes contain semantic cont

researcharxiv-cs-lg
28 May 2026
Safety

Where Rollouts Begin: Low-Load, High-Leverage First-Token Diversification for RLVR

DGX agent

arXiv:2605.28295v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) trains reasoning models without labeled trajectories, relying on grouped rollouts to expose the po

safetyarxiv-cs-ai
28 May 2026
Research

Which Heads Matter for Reasoning? RL-Guided KV Cache Compression

DGX agent

arXiv:2510.08525v3 Announce Type: replace Abstract: Reasoning large language models exhibit complex reasoning behaviors via extended chain-of-thought generation that are highly fragile to information

researcharxiv-cs-cl
28 May 2026
Tutorials

Which Pretraining Paradigm Better Serves Spatial Intelligence? An Empirical Comparison of Vision-Language and Video Generation Models

DGX agent

arXiv:2605.28132v1 Announce Type: new Abstract: Spatial intelligence requires visual representations that capture both semantic objects and geometric structure in the physical world. To support this,

tutorialsarxiv-cs-cv
28 May 2026
← Previous
1…10371038103910401041…1898
Next →