AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
Human
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,019 results
3 Jun 2026

StepFinder: A Temporal Semantic Framework for Failure Attribution in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2606.03467v1 Announce Type: new Abstract: LLM-based multi-agent systems exhibit remarkable collaborative capabilities in complex multi-step tasks. However, these systems are highly sensitive to

Strong signal for local AI on this year's Computex. Big players like NVIDIA and Microsoft are embracing and discussing local AI workloads. D…

Local AiDGX agent

Computex 2024 is expected to feature significant announcements around local AI deployment, with major industry players including NVIDIA and Microsoft actively promoting and discussing on-device AI wor

Strongly Polynomial Time Complexity of Policy Iteration for L_infty Robust MDPs

SafetyDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2601.23229v2 Announce Type: replace Abstract: Markov decision processes (MDPs) are a fundamental model in sequential decision making. Robust MDPs (RMDPs) extend this framework by allowing uncert

Structure-Guided Mixed Masked Pretraining and Spatial Continuity Regularization for Printed Circuit Board Defect Detection

ResearchDGX agent

arXiv:2606.03508v1 Announce Type: new Abstract: Printed circuit board (PCB) defect detection is an essential part of automated optical inspection (AOI); yet it remains challenging in practice because

Structures Facilitate Retrieve, Rerank, and Generate

ResearchDGX agent

arXiv:2606.03247v1 Announce Type: new Abstract: Document-grounded dialogue systems (DGDS) utilize knowledge from external documents to answer domain-specific user questions. Existing solutions typical

Suboptimality bounds for trace-bounded SDPs enable a faster and scalable low-rank SDP solver SDPLR+

ResearchDGX agent

arXiv:2406.10407v3 Announce Type: replace-cross Abstract: Semidefinite programs (SDPs) and their solvers are powerful tools with many applications in machine learning and data science. Designing scala

Suno - a company that trained on 'essentially all music files of reasonable quality that are accessible on the open Internet', and argues it…

SafetyDGX agent

Suno - a company that trained on 'essentially all music files of reasonable quality that are accessible on the open Internet', and argues it does not need to pay to do so - is now valued at $5.4 billi

Supercell, King, and Sybo warn that EU's Digital Fairness Act, requiring pop-ups showing real-world values of virtual currencies, could make games 'unplayable' (Richard Milne/Financial Times)

SafetyDGX agent

Richard Milne / Financial Times: Supercell, King, and Sybo warn that EU's Digital Fairness Act, requiring pop-ups showing real-world values of virtual currencies, could make games “unplayable” — Maker

SVHalluc: Benchmarking Speech-Vision Hallucination in Audio-Visual Large Language Models

Model ReleasesDGX agent

arXiv:2606.02642v1 Announce Type: cross Abstract: Despite the success of audio-visual large-language models (LLMs), they can produce plausible but ungrounded outputs, termed hallucination. Existing be

@swyx and I are curating the AI in GTM track at @aiDotEngineer on June 30. The thing every AI engineer must realize: GTM just became an engi…

AgentsDGX agent

@swyx and I are curating the AI in GTM track at @aiDotEngineer on June 30. The thing every AI engineer must realize: GTM just became an engineering problem. Outbound = agent design. Enrichment = retri

SynCred-Bench: Benchmarking Synthetic Credibility in AI-Generated Visual Misinformation

Model ReleasesDGX agent

arXiv:2606.03348v1 Announce Type: cross Abstract: Recent generative models can now produce visual artifacts with realistic embedded text and layouts, creating a new misinformation threat: synthetic cr

Synthesize and Reward -- Reinforcement Learning for Multi-Step Tool Use in Live Environments

AgentsDGX agent

arXiv:2606.03892v1 Announce Type: cross Abstract: Training LLMs to orchestrate multi-step tool calls is held back by three coupled obstacles: realistic stateful execution environments are costly to bu

Synthetic Hallucinations, Real Gains: Hard Negatives from Frontier Models for FIM Hallucination Mitigation

Model ReleasesDGX agent

arXiv:2606.03130v1 Announce Type: new Abstract: Small open-source code models that power IDE autocomplete still emit hallucinated Fill-in-the-Middle (FIM) completions: syntactically natural calls to m

T2AV-Compass: Towards Unified Evaluation for Text-to-Audio-Video Generation

Model ReleasesDGX agent

arXiv:2512.21094v2 Announce Type: replace Abstract: Text-to-Audio-Video (T2AV) generation aims to synthesize temporally coherent video and semantically synchronized audio from natural language, yet it

TadA-Bench: A Million-Variant Benchmark for Future-Round Discovery Toward Agentic Protein Engineering

Model ReleasesDGX agent

arXiv:2606.02624v1 Announce Type: cross Abstract: AI for scientific discovery is entering an agentic era, where protein-engineering systems are expected to prioritize future wet-lab experiments rather

Taiji: Pareto Optimal Policy Optimization with Semantics-IDs Trade-off for Industrial LLM-Enhanced Recommendation

SafetyDGX agent

arXiv:2606.03866v1 Announce Type: cross Abstract: Scaling recommender systems via large language models (LLMs) has become a prominent trend in the industry. However, aligning the LLM's semantic space

Tailoring Strictly Proper Scoring Rules for Downstream Tasks: An Application to Causal Inference

Local AiDGX agent

arXiv:2606.03332v1 Announce Type: new Abstract: Probabilistic models are typically trained using task-agnostic objectives like log-loss, which can lead to significant errors in downstream estimation.

TalkPlayData 2: An Agentic Synthetic Data Pipeline for Multimodal Conversational Music Recommendation

Model ReleasesDGX agent

arXiv:2509.09685v5 Announce Type: replace-cross Abstract: We present TalkPlayData 2, a synthetic dataset for multimodal conversational music recommendation generated by an agentic data pipeline. In th

Target Updates May Stabilize Linear Q-Learning: Periodic and Soft Dynamics

ResearchDGX agent

arXiv:2606.02645v1 Announce Type: cross Abstract: Periodic target updates in Q-learning and soft target updates in actor-critic methods are empirically well established stabilization mechanisms, but t

TASE: Truncation-Aware Semantic Embeddings for 3D Scene Understanding and Editing

AgentsDGX agent

arXiv:2606.03314v1 Announce Type: new Abstract: High-fidelity semantic 3D scene representations are crucial for numerous applications, including robotics, autonomous driving, and simulation. Beyond th

Template Collapse and Information-Theoretic Limits in Camera rPPG Pulse Morphology Restoration

ResearchDGX agent

arXiv:2606.03802v1 Announce Type: new Abstract: Objective: Consumer face camera remote photoplethysmography (rPPG) enables passive cardiovascular monitoring, but whether single-cycle waveform morpholo

Temporal Action Selection for Action Chunking

SafetyDGX agent

arXiv:2511.04421v2 Announce Type: replace Abstract: Action chunking is a widely adopted approach in Learning from Demonstration (LfD). By modeling multi-step action chunks rather than single-step acti

Terminal Time and Angle-Constrained Nonlinear Intercept Guidance

ResearchDGX agent

arXiv:2606.02872v1 Announce Type: cross Abstract: This paper considers the problem of simultaneously controlling an interceptor's impact time and impact angle using its lateral acceleration as the sol

Test-Time Optimization of Physical Query Plans with LLMs

ResearchDGX agent

arXiv:2602.10387v2 Announce Type: replace-cross Abstract: Traditional query optimization relies on cost-based optimizers that estimate execution cost (e.g., runtime, memory, and I/O) using predefined

Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks

Model ReleasesDGX agent

arXiv:2606.03606v1 Announce Type: cross Abstract: Large language models achieve strong performance on arithmetic reasoning benchmarks, and one common response to arithmetic brittleness is to delegate

Testing Most Influential Sets

ResearchDGX agent

arXiv:2510.20372v4 Announce Type: replace-cross Abstract: Small influential data subsets can dramatically impact model conclusions, with a few data points overturning key findings. While recent work i

Testing the Test: Score-Direction Instability in Class-Split Anomaly Detection

ResearchDGX agent

arXiv:2606.02601v1 Announce Type: new Abstract: Within-dataset class-split evaluation is widely used as a proxy for fully unconditional out-of-distribution anomaly detection. We show that this protoco

TeX-1500: A Paired Real-World LWIR Hyperspectral Dataset and Benchmark for Temperature-Emissivity-Texture Decomposition

Model ReleasesDGX agent

arXiv:2606.03806v1 Announce Type: new Abstract: Temperature-emissivity-texture (TeX) decomposition seeks to recover object heat state, material spectral response, and visible-like geometric texture fr

Text-attributed Graph Condensation via Text Selection and Attribute Matching

ResearchDGX agent

arXiv:2606.03839v1 Announce Type: new Abstract: Text-Attributed Graph (TAG) is an important type of graph structured data, where each node has a text description. TAG models usually train a Graph Neur

Text-to-Image Models Need Less from Text Encoders Than You Think

TutorialsDGX agent

arXiv:2606.03715v1 Announce Type: new Abstract: Text-to-image models rely on text prompts as their primary interface to human intent. Prompts are encoded by a text encoder into embeddings that conditi

TGV-KV: Text-Grounded KV Eviction for Vision-Language Models

SafetyDGX agent

arXiv:2606.03075v1 Announce Type: new Abstract: Vision-Language Models (VLMs) inherit the auto-regressive generation paradigm and cache the keys and values (KV) of all previous tokens to accelerate in

Thanks @hwchase17 @bryonkuchML for the inspiration and the nudge from @MEGAcodePaul - took your feedback seriously. I have developed a boile…

AgentsDGX agent

Thanks @hwchase17 @bryonkuchML for the inspiration and the nudge from @MEGAcodePaul - took your feedback seriously. I have developed a boilerplate that covers the full optimization loop that productio

The Agent's First Day: Benchmarking Learning, Exploration, and Scheduling in the Workplace Scenarios

AgentsDGX agent

arXiv:2601.08173v2 Announce Type: replace Abstract: The rapid evolution of Multi-modal Large Language Models (MLLMs) has advanced workflow automation; however, existing research mainly targets perform

the best agents aren't just built with the best models: they're built with harnesses purpose-built for the task at hand here's a guide on ho…

TutorialsDGX agent

the best agents aren't just built with the best models: they're built with harnesses purpose-built for the task at hand here's a guide on how to build a harness that's really good at feeding the model

The DeepSpeak-Agentic Dataset

Model ReleasesDGX agent

arXiv:2606.03686v1 Announce Type: new Abstract: We present DeepSpeak-Agentic, a dataset of videos comprising over 37 hours of semi-structured conversations between a human and an embodied AI agent. We

The Deliberative Illusion: Diagnosing Factual Attrition and Stance Homogenization in Multi-Agent LLM Deliberation

AgentsDGX agent

arXiv:2606.03032v1 Announce Type: new Abstract: Multi-agent LLM systems often treat consensus as evidence of successful interaction. For deliberative problems, however, reliability depends on whether

The Download: Trump’s new AI order, and smart glasses for warfare

ResearchDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. 5 key points in Trump’s new AI order Less than two weeks after

The Efficiency vs. Accuracy Trade-off: Optimizing RAG-Enhanced LLM Recommender Systems Using Multi-Head Early Exit

ResearchDGX agent

arXiv:2501.02173v2 Announce Type: replace-cross Abstract: The deployment of Large Language Models (LLMs) in recommender systems for predicting Click-Through Rates (CTR) necessitates a delicate balance

The Epi-LLM Framework: probing LLM behavioral priors through epidemiological agent-based models

AgentsDGX agent

arXiv:2606.02867v1 Announce Type: cross Abstract: Human behaviour during epidemics affects infectious disease dynamics, but quantifying this remains deeply challenging. Here we introduce the Epi-LLM f

The Geometry of LLM-as-Judge: Why Inter-LLM Consensus Is Not Human Alignment

Model ReleasesDGX agent

arXiv:2606.03043v1 Announce Type: new Abstract: LMs-as-judges are now standard, yet judges agree strongly with one another while agreeing only weakly with humans. We test whether this reflects shared

The Ghost Annotator: a Framework to Explore Human Label Variation in Content Moderation through Conformal Prediction

SafetyDGX agent

arXiv:2606.02911v1 Announce Type: new Abstract: Current research primarily focuses on model performance, while comparatively less attention has been devoted to uncertainty estimation, particularly in

The Hermes Web Dashboard got a major overhaul: it is now a feature-complete admin panel that you can manage entirely from your browser.

ResearchDGX agent

The Hermes Web Dashboard has been significantly redesigned to function as a comprehensive admin panel with full feature parity, allowing users to manage all operations directly through a web browser i

The Impact of Configuring Agentic AI Coding Tools on Build-vs-Buy Decisions: A Study Protocol

Model ReleasesDGX agent

arXiv:2606.03907v1 Announce Type: cross Abstract: Agentic AI coding tools write code with increasing autonomy and in doing so decide when to import a library and when to implement functionality from s

The Impact of Temporal Granularity on Socio-Demographic Inference from Household Load Profiles

ResearchDGX agent

arXiv:2606.03358v1 Announce Type: new Abstract: Smart meter data can reveal sensitive socio-demographic characteristics of households, raising privacy concerns. While this risk has been demonstrated a

The next chapter in flood resilience: Open sourcing Google’s hydrology framework

ResearchDGX agent

Google has open-sourced its flood forecasting framework, which replicates operational FloodHub model training settings and reflects methodology described in a 2024 Nature paper for global ungauged flo

The Prime Minister says “lessons must be learned” over the death of Henry Nowak. He then goes on to say “there is no two-tier policing” desp…

IndustryDGX agent

The Prime Minister says “lessons must be learned” over the death of Henry Nowak. He then goes on to say “there is no two-tier policing” despite being read a line by Nigel Farage, from the Police racis

THE PROMPT THAT REPLACES YOUR WORKFLOW AUDIT: You are my business transformation strategist. I want you to find the real AI-era business tra…

TutorialsDGX agent

THE PROMPT THAT REPLACES YOUR WORKFLOW AUDIT: You are my business transformation strategist. I want you to find the real AI-era business transformation hiding inside my work by starting from goals, bu

The Reliability Gap in Benchmark Auditing: Distribution Shift and Scale as Failure Modes of Contamination Detection

Model ReleasesDGX agent

arXiv:2606.03305v1 Announce Type: new Abstract: Benchmark contamination, where evaluation examples appear in a model's training data, threatens the validity of LLM assessment. Statistical tools for de

The Ringelmann Effect in Multi-Agent LLM Systems: A Scaling Law for Effective Team Size

Model ReleasesDGX agent

arXiv:2606.02646v1 Announce Type: cross Abstract: Inference-time multi-agent LLM scaling lacks a shared unit: counting nominal agents conflates cost with independent evidence. We derive a two-paramete

The Road Ahead in Autonomous Driving: The KITScenes Multimodal Dataset

Local AiDGX agent

arXiv:2606.02956v1 Announce Type: new Abstract: Existing autonomous driving datasets have enabled major progress, but fall short in sensor fidelity, map completeness, or geographic diversity. We prese

The Shadow Price of Reasoning: Economic Perspective on Optimal Budget Allocation for LLMs

SafetyDGX agent

arXiv:2606.03092v1 Announce Type: new Abstract: Inference-time scaling has emerged as a critical avenue for enhancing Large Language Models' performance, yet real-world deployment is constrained by st

The Shape of Addition: Geometric Structures of Arithmetic in Large Language Models

ResearchDGX agent

arXiv:2606.03645v1 Announce Type: cross Abstract: Large Language Models exhibit paradoxical fragility in fundamental arithmetic, implying a disconnect between internal computation and discrete output.

The Team Behind Deploy: Shipping AI, the DigitalOcean Way

ApplicationsDGX agent

This article profiles the team and processes behind DigitalOcean's Deploy product, which enables users to ship AI applications to production. It likely covers the engineering approach, team structure,

The Unsampled Truth: Psychometrics in SLMs Measure Prompt Artifacts, Not Psychological Constructs

ResearchDGX agent

arXiv:2606.03357v1 Announce Type: cross Abstract: When prompting SLMs for psychometric assessments, researchers assume the outputs reflect semantic reasoning. We evaluate this premise across 13 open-w

The US and other Five Eyes nations warn that China is flooding online job platforms with fake profiles and offers targeting government and military personnel (Greg Miller/Washington Post)

IndustryDGX agent

Greg Miller / Washington Post: The US and other Five Eyes nations warn that China is flooding online job platforms with fake profiles and offers targeting government and military personnel — Nations i

The US data center build-out is falling behind schedule; JP Morgan says 60%+ of data center capacity planned for completion in 2027 isn't yet under construction (Katherine Blunt/Wall Street Journal)

IndustryDGX agent

Katherine Blunt / Wall Street Journal: The US data center build-out is falling behind schedule; JP Morgan says 60%+ of data center capacity planned for completion in 2027 isn't yet under construction

The Violation Situation Pattern: A Knowledge-Graph Pattern for Compliance Violations

ApplicationsDGX agent

arXiv:2606.03326v1 Announce Type: new Abstract: Compliance pipelines detect violations as transient query results and do not keep the violation itself as a persistent graph object with review state, a

The West has created an utterly evil state religion where an accusation of “racism” is the gravest offense that can be committed, even worse…

IndustryDGX agent

The West has created an utterly evil state religion where an accusation of “racism” is the gravest offense that can be committed, even worse than rape or murder! So if police show up at a crime scene

The Word and the Way: Strategies for Domain-Specific BERT Pre-Training in German Medical NLP

Model ReleasesDGX agent

arXiv:2606.03250v1 Announce Type: new Abstract: Digital healthcare generates vast amounts of clinical text that can support AI-assisted applications, yet German biomedical language models remain limit

Theoretical Aspects of Lie Groupoid and Lie Algebroid Equivariant Convolutional Neural Networks

ResearchDGX agent

arXiv:2606.02758v1 Announce Type: cross Abstract: We introduce Lie groupoid equivariant neural networks as a specialization of recently proposed topological category-equivariant neural networks to the

← Previous
1…723724725726727…1517
Next →