AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
1 May 2026

On 5/5 @realDanFu and team will discuss DSV4’s hybrid attention and KV cache efficiency, should be a great session!

Model ReleasesDGX agent

On 5/5 @realDanFu and team will discuss DSV4’s hybrid attention and KV cache efficiency, should be a great session! Join us Tue 5/5: #DeepSeek-V4's hybrid attention + sparse MoE reduces KV cache up to

One week since the launch of GPT-5.5, and it’s already our strongest model launch yet. API revenue is growing more than 2x faster than any p…

Model ReleasesDGX agent

One week since the launch of GPT-5.5, and it’s already our strongest model launch yet. API revenue is growing more than 2x faster than any prior release, while Codex doubled revenue in under seven day

Our CEO @jerryjliu0 in @VentureBeat , on what's actually changing in the LLM stack: 'We've really identified that there's a core set of data…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

Our CEO @jerryjliu0 in @VentureBeat , on what's actually changing in the LLM stack: 'We've really identified that there's a core set of data that has been locked up in all these file format containers

Predictive Multi-Tier Memory Management for KV Cache in Large-Scale GPU Inference

HardwareDGX agent

arXiv:2604.26968v1 Announce Type: cross Abstract: Key-value (KV) cache memory management is the primary bottleneck limiting throughput and cost-efficiency in large-scale GPU inference serving. Current

The new Grok comes in below the latest Chinese open weights models, Grok 4 was at the frontier when released. (& Artificial Analysis: please…

Model ReleasesDGX agent

The new Grok comes in below the latest Chinese open weights models, Grok 4 was at the frontier when released. (& Artificial Analysis: please stop using GDPval-AA which is not a useful test of anything

TopBench: A Benchmark for Implicit Prediction and Reasoning over Tabular Question Answering

Model ReleasesDGX agent

arXiv:2604.28076v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced Table Question Answering, where most queries can be answered by extracting information or simple aggregatio

When training Grok 4.3, we spoke directly with devs and businesses to understand what they actually needed: a model that’s fast, affordable,…

Model ReleasesDGX agent

When training Grok 4.3, we spoke directly with devs and businesses to understand what they actually needed: a model that’s fast, affordable, and great at tool calling. The result is a daily driver tha

30 Apr 2026

A Scaled Three-Vehicle Platooning Platform

SafetyDGX agent

arXiv:2604.25963v1 Announce Type: new Abstract: Vehicle platooning has attracted increasing attention as a promising approach to improve traffic efficiency, energy consumption, and roadway safety thro

A Survey of Process Reward Models: From Outcome Signals to Process Supervisions for Large Language Models

SafetyDGX agent

arXiv:2510.08049v3 Announce Type: replace-cross Abstract: Although Large Language Models (LLMs) exhibit advanced reasoning ability, conventional alignment remains largely dominated by outcome reward m

AGI is GPT-X doing all of the work after you say 'We would love you to throw a party for yourself as a marketing event for OpenAI, so do tha…

ApplicationsDGX agent

This post by Ethan Mollick likely discusses how advanced AI systems like hypothetical future GPT versions could autonomously execute complex, real-world tasks with minimal human direction, using a hum

AMMA: A Multi-Chiplet Memory-Centric Architecture for Low-Latency 1M Context Attention Serving

HardwareDGX agent

arXiv:2604.26103v1 Announce Type: cross Abstract: All current LLM serving systems place the GPU at the center, from production-level attention-FFN disaggregation to NVIDIA's Rubin GPU-LPU heterogeneou

Appian puts reliability at the center of enterprise AI as accuracy gaps frustrate organizations

ApplicationsDGX agent

Enterprise software has no shortage of AI enthusiasm, but the gap between AI potential and production-ready results continues to frustrate organizations investing in enterprise process automation. As

At @sequoia’s AI Ascent last week, @gdb told me something that stuck: in late 2024, AI wrote ~20% of @OpenAI's code. That number is now 80%.…

IndustryDGX agent

At @sequoia’s AI Ascent last week, @gdb told me something that stuck: in late 2024, AI wrote ~20% of @OpenAI's code. That number is now 80%. We also got into why human attention, not compute, is the r

Benchmarking Complex Multimodal Document Processing Pipelines: A Unified Evaluation Framework for Enterprise AI

Model ReleasesDGX agent

arXiv:2604.26382v1 Announce Type: cross Abstract: Most enterprise document AI today is a pipeline. Parse, index, retrieve, generate. Each of those stages has been studied to death on its own -- what's

Checkout DeepAgents deploy here: https://docs.langchain.com/oss/python/deepagents/deploy We're shipping updates every week (almost every day…

ApplicationsDGX agent

LangChain's DeepAgents framework is receiving frequent updates and deployments, with documentation available at their official docs site. The project maintains an active development cycle with updates

Decide less, communicate more: On the construct validity of end-to-end fact-checking in medicine

ResearchDGX agent

arXiv:2506.20876v4 Announce Type: replace Abstract: Technological progress has led to concrete advancements in tasks that were regarded as challenging, such as automatic fact-checking. Interest in ado

Entropy Centroids as Intrinsic Rewards for Test-Time Scaling

Model ReleasesDGX agent

arXiv:2604.26173v1 Announce Type: cross Abstract: An effective way to scale up test-time compute of large language models is to sample multiple responses and then select the best one, as in Grok Heavy

Evergreen: Efficient Claim Verification for Semantic Aggregates

Model ReleasesDGX agent

arXiv:2604.26180v1 Announce Type: cross Abstract: With recent semantic query processing engines, semantic aggregation has become a primitive operator, enabling the reduction of a relation into a natur

FlowS: One-Step Motion Prediction via Local Transport Conditioning

Model ReleasesDGX agent

arXiv:2604.26065v1 Announce Type: new Abstract: Generative motion prediction must satisfy three simultaneous requirements for real-world autonomy: high accuracy, diverse multimodal futures, and strict

ml-intern is fully on mobile now you can launch 8 A100s from your phone. while on the couch. while commuting. wherever I just did this while…

HardwareDGX agent

ml-intern is fully on mobile now you can launch 8 A100s from your phone. while on the couch. while commuting. wherever I just did this while biking. same sessions as your desktop too — start a run on

R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation

SafetyDGX agent

arXiv:2510.08547v2 Announce Type: replace-cross Abstract: Towards the aim of generalized robotic manipulation, spatial generalization is the most fundamental capability that requires the policy to wor

RADIO-ViPE: Online Tightly Coupled Multi-Modal Fusion for Open-Vocabulary Semantic SLAM in Dynamic Environments

Model ReleasesDGX agent

arXiv:2604.26067v1 Announce Type: new Abstract: We present RADIO-ViPE (Reduce All Domains Into One -- Video Pose Engine), an online semantic SLAM system that enables geometry-aware open-vocabulary gro

Student Guides Teacher: Weak-to-Strong Inference via Spectral Orthogonal Exploration

SafetyDGX agent

arXiv:2601.06160v2 Announce Type: replace Abstract: Large Language Models (LLMs) often suffer from ''Reasoning Collapse'' on challenging mathematical reasoning tasks, where stochastic sampling produce

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences

Model ReleasesDGX agent

arXiv:2509.11295v2 Announce Type: replace Abstract: Developing effective prompts demands significant cognitive investment to generate reliable, high-quality responses from Large Language Models (LLMs)

// When to Retrieve During Reasoning // Pay attention to this one, AI devs. (bookmark it) Most RAG systems retrieve once, before the model s…

SafetyDGX agent

// When to Retrieve During Reasoning // Pay attention to this one, AI devs. (bookmark it) Most RAG systems retrieve once, before the model starts reasoning. Large reasoning models like o1 and R1 don't

29 Apr 2026

A million-token context window is not a strategy. 🛑 Our Head of DevRel, @RoieSchwabco , explains why dumping data is killing your RAG perfo…

Model ReleasesDGX agent

A million-token context window is not a strategy. 🛑 Our Head of DevRel, @RoieSchwabco , explains why dumping data is killing your RAG performance: 📍 One needle in a haystack? Easy. 📍 Multiple needles?

BARRED: Synthetic Training of Custom Policy Guardrails via Asymmetric Debate

SafetyDGX agent

arXiv:2604.25203v1 Announce Type: new Abstract: Deploying guardrails for custom policies remains challenging, as generic safety models fail to capture task-specific requirements, while prompting LLMs

CORAL: Adaptive Retrieval Loop for Culturally-Aligned Multilingual RAG

SafetyDGX agent

arXiv:2604.25676v1 Announce Type: new Abstract: Multilingual retrieval-augmented generation (mRAG) is often implemented within a fixed retrieval space, typically via query or document translation or m

Dont Stop Early: Scalable Enterprise Deep Research with Controlled Information Flow and Evidence-Aware Termination

Model ReleasesDGX agent

arXiv:2604.24978v1 Announce Type: new Abstract: Enterprise deep research often fails to produce decision-ready reports due to uneven information coverage, context explosion, and premature stopping. We

Frictive Policy Optimization for LLMs: Epistemic Intervention, Risk-Sensitive Control, and Reflective Alignment

SafetyDGX agent

arXiv:2604.25136v1 Announce Type: new Abstract: We propose Frictive Policy Optimization (FPO), a framework for learning language model policies that regulate not only what to say, but when and how to

GitHub rushed to fix a critical vulnerability in less than six hours

IndustryDGX agent

GitHub employees fixed a critical remote code execution vulnerability in less than six hours last month. Wiz Research used AI models to uncover a vulnerability in GitHub's internal git infrastructure

How RL Unlocks the Aha Moment in Geometric Interleaved Reasoning

SafetyDGX agent

arXiv:2603.01070v2 Announce Type: replace Abstract: Solving complex geometric problems inherently requires interleaved reasoning: a tight alternation between constructing diagrams and performing logic

IMO DeepSeek v4 demonstrated utter confidence and competence by not benchmaxxing, not focusing on some BS final run cost, not even spending …

Model ReleasesDGX agent

IMO DeepSeek v4 demonstrated utter confidence and competence by not benchmaxxing, not focusing on some BS final run cost, not even spending inference-optimal compute. just showed up, demonstrated SOTA

Libra-VLA: Achieving Learning Equilibrium via Asynchronous Coarse-to-Fine Dual-System

SafetyDGX agent

arXiv:2604.24921v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are a promising paradigm for generalist robotic manipulation by grounding high-level semantic instructions into ex

M^3-VQA: A Benchmark for Multimodal, Multi-Entity, Multi-Hop Visual Question Answering

Model ReleasesDGX agent

arXiv:2604.25122v1 Announce Type: new Abstract: We present M^3-VQA, a novel knowledge-based Visual Question Answering (VQA) benchmark, to enhance the evaluation of multimodal large language models (ML

Nemotron 3 Nano Omni: Efficient and Open Multimodal Intelligence

Model ReleasesDGX agent

arXiv:2604.24954v1 Announce Type: cross Abstract: We introduce Nemotron 3 Nano Omni, the latest model in the Nemotron multimodal series and the first to natively support audio inputs alongside text, i

One Perturbation, Two Failure Modes: Probing VLM Safety via Embedding-Guided Typographic Perturbations

Model ReleasesDGX agent

arXiv:2604.25102v1 Announce Type: new Abstract: Typographic prompt injection exploits vision language models' (VLMs) ability to read text rendered in images, posing a growing threat as VLMs power auto

Recreated the app with ml intern to compare. Went faster than in cowork, some things are better (ex direct integration in the desktop reachy…

Model ReleasesDGX agent

Recreated the app with ml intern to compare. Went faster than in cowork, some things are better (ex direct integration in the desktop reachy mini app), some things are worse (had to debut the install)

Toward Multimodal Conversational AI for Age-Related Macular Degeneration

ResearchDGX agent

arXiv:2604.25720v1 Announce Type: cross Abstract: Despite strong performance of deep learning models in retinal disease detection, most systems produce static predictions without clinical reasoning or

What to expect during Atlassian Team ‘26: Join theCUBE May 5-6

IndustryDGX agent

AI-driven workflows are quietly redefining how work actually gets done. As organizations move beyond isolated automation, the emphasis is shifting toward embedding intelligence directly into the flow

28 Apr 2026

Adversary-Free Counterfactual Prediction via Information-Regularized Representations

SafetyDGX agent

arXiv:2510.15479v2 Announce Type: replace Abstract: We study counterfactual prediction under assignment bias and propose a mathematically grounded, information-theoretic approach that removes treatmen

Algorithmic Administration and the EU AI Act: Legal Principles for Public Sector Use of AI

SafetyDGX agent

arXiv:2604.22765v1 Announce Type: cross Abstract: The increasing use of artificial intelligence (AI) by public authorities introduces both opportunities for innovation and significant challenges for t

AnalogRetriever: Learning Cross-Modal Representations for Analog Circuit Retrieval

Model ReleasesDGX agent

arXiv:2604.23195v1 Announce Type: cross Abstract: Analog circuit design relies heavily on reusing existing intellectual property (IP), yet searching across heterogeneous representations such as SPICE

Beyond Match Maximization and Fairness: Retention-Optimized Two-Sided Matching

SafetyDGX agent

arXiv:2602.15752v2 Announce Type: replace Abstract: On two-sided matching platforms such as online dating and recruiting, recommendation algorithms often aim to maximize the total number of matches. H

Bridging the Pose-Semantic Gap: A Cascade Framework for Text-Based Person Anomaly Search

Model ReleasesDGX agent

arXiv:2604.23282v1 Announce Type: new Abstract: Text-based person anomaly search retrieves specific behavioral events from surveillance archives using natural-language queries. Although recent pose-aw

CAP-CoT: Cycle Adversarial Prompt for Improving Chain of Thoughts in LLM Reasoning

SafetyDGX agent

arXiv:2604.23270v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has emerged as a simple and effective way to elicit step-by-step solutions from large language models (LLMs). However,

CAPSULE: Control-Theoretic Action Perturbations for Safe Uncertainty-Aware Reinforcement Learning

SafetyDGX agent

arXiv:2604.23576v1 Announce Type: cross Abstract: Ensuring safe exploration in high-dimensional systems with unknown dynamics remains a significant challenge. Existing safe reinforcement learning meth

Caries DETR: Tooth Structure-aware Prior and Lesion-aware Dynamic Loss Refinement for DETR Based Caries Detection

ResearchDGX agent

arXiv:2604.23718v1 Announce Type: new Abstract: As dental caries appear as subtle, low-contrast lesions in intraoral imaging, existing deep learning models face significant challenges in the early det

CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era

Model ReleasesDGX agent

arXiv:2602.23452v2 Announce Type: replace Abstract: Scientific research relies on accurate citation for attribution and integrity, yet large language models (LLMs) introduce a new risk: fabricated ref

CorpusQA: A 10 Million Token Benchmark for Corpus-Level Analysis and Reasoning

Model ReleasesDGX agent

arXiv:2601.14952v2 Announce Type: replace-cross Abstract: While large language models now handle million-token contexts, their capacity for reasoning across entire document repositories remains largel

Discovering Failure Modes in Vision-Language Models using RL

SafetyDGX agent

arXiv:2604.04733v2 Announce Type: replace-cross Abstract: Vision-language Models (VLMs), despite achieving strong performance on multimodal benchmarks, often misinterpret straightforward visual concep

Enterprises are not running out of AI ambition — they are running out of time to act on it

Model ReleasesDGX agent

Enterprise AI transformation has a new home: the boardroom. Across financial services, CEOs are now demanding results, not more roadmaps. As Google Cloud Next 2026 packed Las Vegas with announcements

Evaluating Language Models' Evaluations of Games

SafetyDGX agent

arXiv:2510.10930v2 Announce Type: replace-cross Abstract: Reasoning is not just about solving problems -- it is also about evaluating which problems are worth solving at all. Evaluations of artificial

Excited to support @NVIDIA Nemotron 3 Nano Omni, now available on Fireworks. It's the first open model that handles vision, audio, video, an…

Model ReleasesDGX agent

Excited to support @NVIDIA Nemotron 3 Nano Omni, now available on Fireworks. It's the first open model that handles vision, audio, video, and text in a single inference loop. Built for multimodal sub-

Extending Precipitation Nowcasting Horizons via Spectral Fusion of Radar Observations and Foundation Model Priors

SafetyDGX agent

arXiv:2603.21768v3 Announce Type: replace-cross Abstract: Precipitation nowcasting is critical for disaster mitigation and aviation safety. However, radar-only models frequently suffer from a lack of

First open-weight model from @poolsideai! Apache license, and available on Ollama to try. 👇👇👇 model page

Model ReleasesDGX agent

First open-weight model from @poolsideai! Apache license, and available on Ollama to try. 👇👇👇 model page Today we’re releasing Laguna XS.2, Poolside’s first open-weight model. It’s a 33B total / 3B ac

GeoFunFlow-3D: A Physics-Guided Generative Flow Matching Framework for High-Fidelity 3D Aerodynamic Inference over Complex Geometries

Local AiDGX agent

arXiv:2604.23350v1 Announce Type: cross Abstract: Deep generative models and neural operators have demonstrated significant potential for 3D aerodynamic inference. However, they often face inherent ch

GLM 5.1 from @Zai_org is now available on @FireworksAI_HQ Training Platform across the Managed and Training API workflows. Try SFT and DPO w…

Model ReleasesDGX agent

GLM 5.1 from @Zai_org is now available on @FireworksAI_HQ Training Platform across the Managed and Training API workflows. Try SFT and DPO with smart defaults or your own custom loss function with a 2

Green Shielding: A User-Centric Approach Towards Trustworthy AI

Model ReleasesDGX agent

arXiv:2604.24700v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed, yet their outputs can be highly sensitive to routine, non-adversarial variation in how users p

Grounding Before Generalizing: How AI Differs from Humans in Causal Transfer

ResearchDGX agent

arXiv:2604.24062v1 Announce Type: new Abstract: Extracting abstract causal structures and applying them to novel situations is a hallmark of human intelligence. While Large Language Models (LLMs) and

← Previous
1…282283284285286…294
Next →