AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,607 results
Model Releases

Benchmarking Complex Multimodal Document Processing Pipelines: A Unified Evaluation Framework for Enterprise AI

DGX agent

arXiv:2604.26382v1 Announce Type: cross Abstract: Most enterprise document AI today is a pipeline. Parse, index, retrieve, generate. Each of those stages has been studied to death on its own -- what's

model-releasesarxiv-cs-ai
30 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

Checkout DeepAgents deploy here: https://docs.langchain.com/oss/python/deepagents/deploy We're shipping updates every week (almost every day…

DGX agent

LangChain's DeepAgents framework is receiving frequent updates and deployments, with documentation available at their official docs site. The project maintains an active development cycle with updates

applicationsharrison-chase--x
30 Apr 2026
Research

Decide less, communicate more: On the construct validity of end-to-end fact-checking in medicine

DGX agent

arXiv:2506.20876v4 Announce Type: replace Abstract: Technological progress has led to concrete advancements in tasks that were regarded as challenging, such as automatic fact-checking. Interest in ado

researcharxiv-cs-cl
30 Apr 2026
Model Releases

Entropy Centroids as Intrinsic Rewards for Test-Time Scaling

DGX agent

arXiv:2604.26173v1 Announce Type: cross Abstract: An effective way to scale up test-time compute of large language models is to sample multiple responses and then select the best one, as in Grok Heavy

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Evergreen: Efficient Claim Verification for Semantic Aggregates

DGX agent

arXiv:2604.26180v1 Announce Type: cross Abstract: With recent semantic query processing engines, semantic aggregation has become a primitive operator, enabling the reduction of a relation into a natur

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

FlowS: One-Step Motion Prediction via Local Transport Conditioning

DGX agent

arXiv:2604.26065v1 Announce Type: new Abstract: Generative motion prediction must satisfy three simultaneous requirements for real-world autonomy: high accuracy, diverse multimodal futures, and strict

model-releasesarxiv-cs-ro
30 Apr 2026
Hardware

ml-intern is fully on mobile now you can launch 8 A100s from your phone. while on the couch. while commuting. wherever I just did this while…

DGX agent

ml-intern is fully on mobile now you can launch 8 A100s from your phone. while on the couch. while commuting. wherever I just did this while biking. same sessions as your desktop too — start a run on

hardwareclem-delangue--x
30 Apr 2026
Safety

R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation

DGX agent

arXiv:2510.08547v2 Announce Type: replace-cross Abstract: Towards the aim of generalized robotic manipulation, spatial generalization is the most fundamental capability that requires the policy to wor

safetyarxiv-cs-cv
30 Apr 2026
Model Releases

RADIO-ViPE: Online Tightly Coupled Multi-Modal Fusion for Open-Vocabulary Semantic SLAM in Dynamic Environments

DGX agent

arXiv:2604.26067v1 Announce Type: new Abstract: We present RADIO-ViPE (Reduce All Domains Into One -- Video Pose Engine), an online semantic SLAM system that enables geometry-aware open-vocabulary gro

model-releasesarxiv-cs-cv
30 Apr 2026
Safety

Student Guides Teacher: Weak-to-Strong Inference via Spectral Orthogonal Exploration

DGX agent

arXiv:2601.06160v2 Announce Type: replace Abstract: Large Language Models (LLMs) often suffer from ''Reasoning Collapse'' on challenging mathematical reasoning tasks, where stochastic sampling produce

safetyarxiv-cs-ai
30 Apr 2026
Model Releases

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences

DGX agent

arXiv:2509.11295v2 Announce Type: replace Abstract: Developing effective prompts demands significant cognitive investment to generate reliable, high-quality responses from Large Language Models (LLMs)

model-releasesarxiv-cs-cl
30 Apr 2026
Safety

// When to Retrieve During Reasoning // Pay attention to this one, AI devs. (bookmark it) Most RAG systems retrieve once, before the model s…

DGX agent

// When to Retrieve During Reasoning // Pay attention to this one, AI devs. (bookmark it) Most RAG systems retrieve once, before the model starts reasoning. Large reasoning models like o1 and R1 don't

safetydair-ai--x
30 Apr 2026
Model Releases

A million-token context window is not a strategy. 🛑 Our Head of DevRel, @RoieSchwabco , explains why dumping data is killing your RAG perfo…

DGX agent

A million-token context window is not a strategy. 🛑 Our Head of DevRel, @RoieSchwabco , explains why dumping data is killing your RAG performance: 📍 One needle in a haystack? Easy. 📍 Multiple needles?

model-releasespinecone--x
29 Apr 2026
Safety

BARRED: Synthetic Training of Custom Policy Guardrails via Asymmetric Debate

DGX agent

arXiv:2604.25203v1 Announce Type: new Abstract: Deploying guardrails for custom policies remains challenging, as generic safety models fail to capture task-specific requirements, while prompting LLMs

safetyarxiv-cs-cl
29 Apr 2026
Safety

CORAL: Adaptive Retrieval Loop for Culturally-Aligned Multilingual RAG

DGX agent

arXiv:2604.25676v1 Announce Type: new Abstract: Multilingual retrieval-augmented generation (mRAG) is often implemented within a fixed retrieval space, typically via query or document translation or m

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

Dont Stop Early: Scalable Enterprise Deep Research with Controlled Information Flow and Evidence-Aware Termination

DGX agent

arXiv:2604.24978v1 Announce Type: new Abstract: Enterprise deep research often fails to produce decision-ready reports due to uneven information coverage, context explosion, and premature stopping. We

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

Frictive Policy Optimization for LLMs: Epistemic Intervention, Risk-Sensitive Control, and Reflective Alignment

DGX agent

arXiv:2604.25136v1 Announce Type: new Abstract: We propose Frictive Policy Optimization (FPO), a framework for learning language model policies that regulate not only what to say, but when and how to

safetyarxiv-cs-cl
29 Apr 2026
Industry

GitHub rushed to fix a critical vulnerability in less than six hours

DGX agent

GitHub employees fixed a critical remote code execution vulnerability in less than six hours last month. Wiz Research used AI models to uncover a vulnerability in GitHub's internal git infrastructure

industrythe-verge-ai
29 Apr 2026
Safety

How RL Unlocks the Aha Moment in Geometric Interleaved Reasoning

DGX agent

arXiv:2603.01070v2 Announce Type: replace Abstract: Solving complex geometric problems inherently requires interleaved reasoning: a tight alternation between constructing diagrams and performing logic

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

IMO DeepSeek v4 demonstrated utter confidence and competence by not benchmaxxing, not focusing on some BS final run cost, not even spending …

DGX agent

IMO DeepSeek v4 demonstrated utter confidence and competence by not benchmaxxing, not focusing on some BS final run cost, not even spending inference-optimal compute. just showed up, demonstrated SOTA

model-releasesswyx--x
29 Apr 2026
Safety

Libra-VLA: Achieving Learning Equilibrium via Asynchronous Coarse-to-Fine Dual-System

DGX agent

arXiv:2604.24921v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are a promising paradigm for generalist robotic manipulation by grounding high-level semantic instructions into ex

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

M^3-VQA: A Benchmark for Multimodal, Multi-Entity, Multi-Hop Visual Question Answering

DGX agent

arXiv:2604.25122v1 Announce Type: new Abstract: We present M^3-VQA, a novel knowledge-based Visual Question Answering (VQA) benchmark, to enhance the evaluation of multimodal large language models (ML

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Nemotron 3 Nano Omni: Efficient and Open Multimodal Intelligence

DGX agent

arXiv:2604.24954v1 Announce Type: cross Abstract: We introduce Nemotron 3 Nano Omni, the latest model in the Nemotron multimodal series and the first to natively support audio inputs alongside text, i

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

One Perturbation, Two Failure Modes: Probing VLM Safety via Embedding-Guided Typographic Perturbations

DGX agent

arXiv:2604.25102v1 Announce Type: new Abstract: Typographic prompt injection exploits vision language models' (VLMs) ability to read text rendered in images, posing a growing threat as VLMs power auto

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Recreated the app with ml intern to compare. Went faster than in cowork, some things are better (ex direct integration in the desktop reachy…

DGX agent

Recreated the app with ml intern to compare. Went faster than in cowork, some things are better (ex direct integration in the desktop reachy mini app), some things are worse (had to debut the install)

model-releasesclem-delangue--x
29 Apr 2026
Research

Toward Multimodal Conversational AI for Age-Related Macular Degeneration

DGX agent

arXiv:2604.25720v1 Announce Type: cross Abstract: Despite strong performance of deep learning models in retinal disease detection, most systems produce static predictions without clinical reasoning or

researcharxiv-cs-cl
29 Apr 2026
Industry

What to expect during Atlassian Team ‘26: Join theCUBE May 5-6

DGX agent

AI-driven workflows are quietly redefining how work actually gets done. As organizations move beyond isolated automation, the emphasis is shifting toward embedding intelligence directly into the flow

industrysiliconangle
29 Apr 2026
Safety

Adversary-Free Counterfactual Prediction via Information-Regularized Representations

DGX agent

arXiv:2510.15479v2 Announce Type: replace Abstract: We study counterfactual prediction under assignment bias and propose a mathematically grounded, information-theoretic approach that removes treatmen

safetyarxiv-cs-lg
28 Apr 2026
Safety

Algorithmic Administration and the EU AI Act: Legal Principles for Public Sector Use of AI

DGX agent

arXiv:2604.22765v1 Announce Type: cross Abstract: The increasing use of artificial intelligence (AI) by public authorities introduces both opportunities for innovation and significant challenges for t

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

AnalogRetriever: Learning Cross-Modal Representations for Analog Circuit Retrieval

DGX agent

arXiv:2604.23195v1 Announce Type: cross Abstract: Analog circuit design relies heavily on reusing existing intellectual property (IP), yet searching across heterogeneous representations such as SPICE

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Beyond Match Maximization and Fairness: Retention-Optimized Two-Sided Matching

DGX agent

arXiv:2602.15752v2 Announce Type: replace Abstract: On two-sided matching platforms such as online dating and recruiting, recommendation algorithms often aim to maximize the total number of matches. H

safetyarxiv-cs-lg
28 Apr 2026
Model Releases

Bridging the Pose-Semantic Gap: A Cascade Framework for Text-Based Person Anomaly Search

DGX agent

arXiv:2604.23282v1 Announce Type: new Abstract: Text-based person anomaly search retrieves specific behavioral events from surveillance archives using natural-language queries. Although recent pose-aw

model-releasesarxiv-cs-cv
28 Apr 2026
Safety

CAP-CoT: Cycle Adversarial Prompt for Improving Chain of Thoughts in LLM Reasoning

DGX agent

arXiv:2604.23270v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has emerged as a simple and effective way to elicit step-by-step solutions from large language models (LLMs). However,

safetyarxiv-cs-ai
28 Apr 2026
Safety

CAPSULE: Control-Theoretic Action Perturbations for Safe Uncertainty-Aware Reinforcement Learning

DGX agent

arXiv:2604.23576v1 Announce Type: cross Abstract: Ensuring safe exploration in high-dimensional systems with unknown dynamics remains a significant challenge. Existing safe reinforcement learning meth

safetyarxiv-cs-ai
28 Apr 2026
Research

Caries DETR: Tooth Structure-aware Prior and Lesion-aware Dynamic Loss Refinement for DETR Based Caries Detection

DGX agent

arXiv:2604.23718v1 Announce Type: new Abstract: As dental caries appear as subtle, low-contrast lesions in intraoral imaging, existing deep learning models face significant challenges in the early det

researcharxiv-cs-cv
28 Apr 2026
Model Releases

CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era

DGX agent

arXiv:2602.23452v2 Announce Type: replace Abstract: Scientific research relies on accurate citation for attribution and integrity, yet large language models (LLMs) introduce a new risk: fabricated ref

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

CorpusQA: A 10 Million Token Benchmark for Corpus-Level Analysis and Reasoning

DGX agent

arXiv:2601.14952v2 Announce Type: replace-cross Abstract: While large language models now handle million-token contexts, their capacity for reasoning across entire document repositories remains largel

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Discovering Failure Modes in Vision-Language Models using RL

DGX agent

arXiv:2604.04733v2 Announce Type: replace-cross Abstract: Vision-language Models (VLMs), despite achieving strong performance on multimodal benchmarks, often misinterpret straightforward visual concep

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Enterprises are not running out of AI ambition — they are running out of time to act on it

DGX agent

Enterprise AI transformation has a new home: the boardroom. Across financial services, CEOs are now demanding results, not more roadmaps. As Google Cloud Next 2026 packed Las Vegas with announcements

model-releasessiliconangle
28 Apr 2026
Safety

Evaluating Language Models' Evaluations of Games

DGX agent

arXiv:2510.10930v2 Announce Type: replace-cross Abstract: Reasoning is not just about solving problems -- it is also about evaluating which problems are worth solving at all. Evaluations of artificial

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Excited to support @NVIDIA Nemotron 3 Nano Omni, now available on Fireworks. It's the first open model that handles vision, audio, video, an…

DGX agent

Excited to support @NVIDIA Nemotron 3 Nano Omni, now available on Fireworks. It's the first open model that handles vision, audio, video, and text in a single inference loop. Built for multimodal sub-

model-releasesfireworks-ai--x
28 Apr 2026
Safety

Extending Precipitation Nowcasting Horizons via Spectral Fusion of Radar Observations and Foundation Model Priors

DGX agent

arXiv:2603.21768v3 Announce Type: replace-cross Abstract: Precipitation nowcasting is critical for disaster mitigation and aviation safety. However, radar-only models frequently suffer from a lack of

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

First open-weight model from @poolsideai! Apache license, and available on Ollama to try. 👇👇👇 model page

DGX agent

First open-weight model from @poolsideai! Apache license, and available on Ollama to try. 👇👇👇 model page Today we’re releasing Laguna XS.2, Poolside’s first open-weight model. It’s a 33B total / 3B ac

model-releasesollama--x
28 Apr 2026
Local Ai

GeoFunFlow-3D: A Physics-Guided Generative Flow Matching Framework for High-Fidelity 3D Aerodynamic Inference over Complex Geometries

DGX agent

arXiv:2604.23350v1 Announce Type: cross Abstract: Deep generative models and neural operators have demonstrated significant potential for 3D aerodynamic inference. However, they often face inherent ch

local-aiarxiv-cs-lg
28 Apr 2026
Model Releases

GLM 5.1 from @Zai_org is now available on @FireworksAI_HQ Training Platform across the Managed and Training API workflows. Try SFT and DPO w…

DGX agent

GLM 5.1 from @Zai_org is now available on @FireworksAI_HQ Training Platform across the Managed and Training API workflows. Try SFT and DPO with smart defaults or your own custom loss function with a 2

model-releasesfireworks-ai--x
28 Apr 2026
Model Releases

Green Shielding: A User-Centric Approach Towards Trustworthy AI

DGX agent

arXiv:2604.24700v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed, yet their outputs can be highly sensitive to routine, non-adversarial variation in how users p

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Grounding Before Generalizing: How AI Differs from Humans in Causal Transfer

DGX agent

arXiv:2604.24062v1 Announce Type: new Abstract: Extracting abstract causal structures and applying them to novel situations is a hallmark of human intelligence. While Large Language Models (LLMs) and

researcharxiv-cs-ai
28 Apr 2026
Safety

Institutions for the Post-Scarcity of Judgment

DGX agent

arXiv:2604.22966v1 Announce Type: cross Abstract: Each major technological revolution inverts a particular scarcity and rebuilds institutions around the shift. The near-consensus diagnosis of the AI r

safetyarxiv-cs-ai
28 Apr 2026
← Previous
1…353354355356357…367
Next →