AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,682
  • Agents7,811
  • Applications5,571
  • Concepts5
  • Hardware1,953
  • Industry6,231
  • Local Ai5,135
  • Model Releases25,010
  • Research20,929
  • Safety13,841
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,682
  • Agents7,811
  • Applications5,571
  • Concepts5
  • Hardware1,953
  • Industry6,231
  • Local Ai5,135
  • Model Releases25,010
  • Research20,929
  • Safety13,841
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
91,682Total entries
1Added by human
91,681Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,681 results
Model Releases

SynCred-Bench: Benchmarking Synthetic Credibility in AI-Generated Visual Misinformation

DGX agent

arXiv:2606.03348v1 Announce Type: cross Abstract: Recent generative models can now produce visual artifacts with realistic embedded text and layouts, creating a new misinformation threat: synthetic cr

model-releasesarxiv-cs-ai
3 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Synthesize and Reward -- Reinforcement Learning for Multi-Step Tool Use in Live Environments

DGX agent

arXiv:2606.03892v1 Announce Type: cross Abstract: Training LLMs to orchestrate multi-step tool calls is held back by three coupled obstacles: realistic stateful execution environments are costly to bu

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

Synthetic Hallucinations, Real Gains: Hard Negatives from Frontier Models for FIM Hallucination Mitigation

DGX agent

arXiv:2606.03130v1 Announce Type: new Abstract: Small open-source code models that power IDE autocomplete still emit hallucinated Fill-in-the-Middle (FIM) completions: syntactically natural calls to m

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

T2AV-Compass: Towards Unified Evaluation for Text-to-Audio-Video Generation

DGX agent

arXiv:2512.21094v2 Announce Type: replace Abstract: Text-to-Audio-Video (T2AV) generation aims to synthesize temporally coherent video and semantically synchronized audio from natural language, yet it

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

TadA-Bench: A Million-Variant Benchmark for Future-Round Discovery Toward Agentic Protein Engineering

DGX agent

arXiv:2606.02624v1 Announce Type: cross Abstract: AI for scientific discovery is entering an agentic era, where protein-engineering systems are expected to prioritize future wet-lab experiments rather

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Taiji: Pareto Optimal Policy Optimization with Semantics-IDs Trade-off for Industrial LLM-Enhanced Recommendation

DGX agent

arXiv:2606.03866v1 Announce Type: cross Abstract: Scaling recommender systems via large language models (LLMs) has become a prominent trend in the industry. However, aligning the LLM's semantic space

safetyarxiv-cs-ai
3 Jun 2026
Local Ai

Tailoring Strictly Proper Scoring Rules for Downstream Tasks: An Application to Causal Inference

DGX agent

arXiv:2606.03332v1 Announce Type: new Abstract: Probabilistic models are typically trained using task-agnostic objectives like log-loss, which can lead to significant errors in downstream estimation.

local-aiarxiv-cs-lg
3 Jun 2026
Model Releases

TalkPlayData 2: An Agentic Synthetic Data Pipeline for Multimodal Conversational Music Recommendation

DGX agent

arXiv:2509.09685v5 Announce Type: replace-cross Abstract: We present TalkPlayData 2, a synthetic dataset for multimodal conversational music recommendation generated by an agentic data pipeline. In th

model-releasesarxiv-cs-ai
3 Jun 2026
Research

Target Updates May Stabilize Linear Q-Learning: Periodic and Soft Dynamics

DGX agent

arXiv:2606.02645v1 Announce Type: cross Abstract: Periodic target updates in Q-learning and soft target updates in actor-critic methods are empirically well established stabilization mechanisms, but t

researcharxiv-cs-ai
3 Jun 2026
Agents

TASE: Truncation-Aware Semantic Embeddings for 3D Scene Understanding and Editing

DGX agent

arXiv:2606.03314v1 Announce Type: new Abstract: High-fidelity semantic 3D scene representations are crucial for numerous applications, including robotics, autonomous driving, and simulation. Beyond th

agentsarxiv-cs-cv
3 Jun 2026
Research

Template Collapse and Information-Theoretic Limits in Camera rPPG Pulse Morphology Restoration

DGX agent

arXiv:2606.03802v1 Announce Type: new Abstract: Objective: Consumer face camera remote photoplethysmography (rPPG) enables passive cardiovascular monitoring, but whether single-cycle waveform morpholo

researcharxiv-cs-cv
3 Jun 2026
Safety

Temporal Action Selection for Action Chunking

DGX agent

arXiv:2511.04421v2 Announce Type: replace Abstract: Action chunking is a widely adopted approach in Learning from Demonstration (LfD). By modeling multi-step action chunks rather than single-step acti

safetyarxiv-cs-ro
3 Jun 2026
Research

Terminal Time and Angle-Constrained Nonlinear Intercept Guidance

DGX agent

arXiv:2606.02872v1 Announce Type: cross Abstract: This paper considers the problem of simultaneously controlling an interceptor's impact time and impact angle using its lateral acceleration as the sol

researcharxiv-cs-ro
3 Jun 2026
Research

Test-Time Optimization of Physical Query Plans with LLMs

DGX agent

arXiv:2602.10387v2 Announce Type: replace-cross Abstract: Traditional query optimization relies on cost-based optimizers that estimate execution cost (e.g., runtime, memory, and I/O) using predefined

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks

DGX agent

arXiv:2606.03606v1 Announce Type: cross Abstract: Large language models achieve strong performance on arithmetic reasoning benchmarks, and one common response to arithmetic brittleness is to delegate

model-releasesarxiv-cs-ai
3 Jun 2026
Research

Testing Most Influential Sets

DGX agent

arXiv:2510.20372v4 Announce Type: replace-cross Abstract: Small influential data subsets can dramatically impact model conclusions, with a few data points overturning key findings. While recent work i

researcharxiv-cs-lg
3 Jun 2026
Research

Testing the Test: Score-Direction Instability in Class-Split Anomaly Detection

DGX agent

arXiv:2606.02601v1 Announce Type: new Abstract: Within-dataset class-split evaluation is widely used as a proxy for fully unconditional out-of-distribution anomaly detection. We show that this protoco

researcharxiv-cs-lg
3 Jun 2026
Model Releases

TeX-1500: A Paired Real-World LWIR Hyperspectral Dataset and Benchmark for Temperature-Emissivity-Texture Decomposition

DGX agent

arXiv:2606.03806v1 Announce Type: new Abstract: Temperature-emissivity-texture (TeX) decomposition seeks to recover object heat state, material spectral response, and visible-like geometric texture fr

model-releasesarxiv-cs-cv
3 Jun 2026
Research

Text-attributed Graph Condensation via Text Selection and Attribute Matching

DGX agent

arXiv:2606.03839v1 Announce Type: new Abstract: Text-Attributed Graph (TAG) is an important type of graph structured data, where each node has a text description. TAG models usually train a Graph Neur

researcharxiv-cs-lg
3 Jun 2026
Tutorials

Text-to-Image Models Need Less from Text Encoders Than You Think

DGX agent

arXiv:2606.03715v1 Announce Type: new Abstract: Text-to-image models rely on text prompts as their primary interface to human intent. Prompts are encoded by a text encoder into embeddings that conditi

tutorialsarxiv-cs-cv
3 Jun 2026
Safety

TGV-KV: Text-Grounded KV Eviction for Vision-Language Models

DGX agent

arXiv:2606.03075v1 Announce Type: new Abstract: Vision-Language Models (VLMs) inherit the auto-regressive generation paradigm and cache the keys and values (KV) of all previous tokens to accelerate in

safetyarxiv-cs-cv
3 Jun 2026
Agents

Thanks @hwchase17 @bryonkuchML for the inspiration and the nudge from @MEGAcodePaul - took your feedback seriously. I have developed a boile…

DGX agent

Thanks @hwchase17 @bryonkuchML for the inspiration and the nudge from @MEGAcodePaul - took your feedback seriously. I have developed a boilerplate that covers the full optimization loop that productio

agentsharrison-chase--x
3 Jun 2026
Agents

The Agent's First Day: Benchmarking Learning, Exploration, and Scheduling in the Workplace Scenarios

DGX agent

arXiv:2601.08173v2 Announce Type: replace Abstract: The rapid evolution of Multi-modal Large Language Models (MLLMs) has advanced workflow automation; however, existing research mainly targets perform

agentsarxiv-cs-ai
3 Jun 2026
Tutorials

the best agents aren't just built with the best models: they're built with harnesses purpose-built for the task at hand here's a guide on ho…

DGX agent

the best agents aren't just built with the best models: they're built with harnesses purpose-built for the task at hand here's a guide on how to build a harness that's really good at feeding the model

tutorialsharrison-chase--x
3 Jun 2026
Model Releases

The DeepSpeak-Agentic Dataset

DGX agent

arXiv:2606.03686v1 Announce Type: new Abstract: We present DeepSpeak-Agentic, a dataset of videos comprising over 37 hours of semi-structured conversations between a human and an embodied AI agent. We

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

The Deliberative Illusion: Diagnosing Factual Attrition and Stance Homogenization in Multi-Agent LLM Deliberation

DGX agent

arXiv:2606.03032v1 Announce Type: new Abstract: Multi-agent LLM systems often treat consensus as evidence of successful interaction. For deliberative problems, however, reliability depends on whether

agentsarxiv-cs-cl
3 Jun 2026
Research

The Download: Trump’s new AI order, and smart glasses for warfare

DGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. 5 key points in Trump’s new AI order Less than two weeks after

researchmit-tech-review
3 Jun 2026
Research

The Efficiency vs. Accuracy Trade-off: Optimizing RAG-Enhanced LLM Recommender Systems Using Multi-Head Early Exit

DGX agent

arXiv:2501.02173v2 Announce Type: replace-cross Abstract: The deployment of Large Language Models (LLMs) in recommender systems for predicting Click-Through Rates (CTR) necessitates a delicate balance

researcharxiv-cs-lg
3 Jun 2026
Agents

The Epi-LLM Framework: probing LLM behavioral priors through epidemiological agent-based models

DGX agent

arXiv:2606.02867v1 Announce Type: cross Abstract: Human behaviour during epidemics affects infectious disease dynamics, but quantifying this remains deeply challenging. Here we introduce the Epi-LLM f

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

The Geometry of LLM-as-Judge: Why Inter-LLM Consensus Is Not Human Alignment

DGX agent

arXiv:2606.03043v1 Announce Type: new Abstract: LMs-as-judges are now standard, yet judges agree strongly with one another while agreeing only weakly with humans. We test whether this reflects shared

model-releasesarxiv-cs-cl
3 Jun 2026
Safety

The Ghost Annotator: a Framework to Explore Human Label Variation in Content Moderation through Conformal Prediction

DGX agent

arXiv:2606.02911v1 Announce Type: new Abstract: Current research primarily focuses on model performance, while comparatively less attention has been devoted to uncertainty estimation, particularly in

safetyarxiv-cs-cl
3 Jun 2026
Research

The Hermes Web Dashboard got a major overhaul: it is now a feature-complete admin panel that you can manage entirely from your browser.

DGX agent

The Hermes Web Dashboard has been significantly redesigned to function as a comprehensive admin panel with full feature parity, allowing users to manage all operations directly through a web browser i

researchnous-research--x
3 Jun 2026
Model Releases

The Impact of Configuring Agentic AI Coding Tools on Build-vs-Buy Decisions: A Study Protocol

DGX agent

arXiv:2606.03907v1 Announce Type: cross Abstract: Agentic AI coding tools write code with increasing autonomy and in doing so decide when to import a library and when to implement functionality from s

model-releasesarxiv-cs-ai
3 Jun 2026
Research

The Impact of Temporal Granularity on Socio-Demographic Inference from Household Load Profiles

DGX agent

arXiv:2606.03358v1 Announce Type: new Abstract: Smart meter data can reveal sensitive socio-demographic characteristics of households, raising privacy concerns. While this risk has been demonstrated a

researcharxiv-cs-lg
3 Jun 2026
Research

The next chapter in flood resilience: Open sourcing Google’s hydrology framework

DGX agent

Google has open-sourced its flood forecasting framework, which replicates operational FloodHub model training settings and reflects methodology described in a 2024 Nature paper for global ungauged flo

researchgoogle-research
3 Jun 2026
Industry

The Prime Minister says “lessons must be learned” over the death of Henry Nowak. He then goes on to say “there is no two-tier policing” desp…

DGX agent

The Prime Minister says “lessons must be learned” over the death of Henry Nowak. He then goes on to say “there is no two-tier policing” despite being read a line by Nigel Farage, from the Police racis

industryelon-musk--x
3 Jun 2026
Tutorials

THE PROMPT THAT REPLACES YOUR WORKFLOW AUDIT: You are my business transformation strategist. I want you to find the real AI-era business tra…

DGX agent

THE PROMPT THAT REPLACES YOUR WORKFLOW AUDIT: You are my business transformation strategist. I want you to find the real AI-era business transformation hiding inside my work by starting from goals, bu

tutorialsallie-k--miller--x
3 Jun 2026
Model Releases

The Reliability Gap in Benchmark Auditing: Distribution Shift and Scale as Failure Modes of Contamination Detection

DGX agent

arXiv:2606.03305v1 Announce Type: new Abstract: Benchmark contamination, where evaluation examples appear in a model's training data, threatens the validity of LLM assessment. Statistical tools for de

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

The Ringelmann Effect in Multi-Agent LLM Systems: A Scaling Law for Effective Team Size

DGX agent

arXiv:2606.02646v1 Announce Type: cross Abstract: Inference-time multi-agent LLM scaling lacks a shared unit: counting nominal agents conflates cost with independent evidence. We derive a two-paramete

model-releasesarxiv-cs-ai
3 Jun 2026
Local Ai

The Road Ahead in Autonomous Driving: The KITScenes Multimodal Dataset

DGX agent

arXiv:2606.02956v1 Announce Type: new Abstract: Existing autonomous driving datasets have enabled major progress, but fall short in sensor fidelity, map completeness, or geographic diversity. We prese

local-aiarxiv-cs-cv
3 Jun 2026
Safety

The Shadow Price of Reasoning: Economic Perspective on Optimal Budget Allocation for LLMs

DGX agent

arXiv:2606.03092v1 Announce Type: new Abstract: Inference-time scaling has emerged as a critical avenue for enhancing Large Language Models' performance, yet real-world deployment is constrained by st

safetyarxiv-cs-ai
3 Jun 2026
Research

The Shape of Addition: Geometric Structures of Arithmetic in Large Language Models

DGX agent

arXiv:2606.03645v1 Announce Type: cross Abstract: Large Language Models exhibit paradoxical fragility in fundamental arithmetic, implying a disconnect between internal computation and discrete output.

researcharxiv-cs-ai
3 Jun 2026
Applications

The Team Behind Deploy: Shipping AI, the DigitalOcean Way

DGX agent

This article profiles the team and processes behind DigitalOcean's Deploy product, which enables users to ship AI applications to production. It likely covers the engineering approach, team structure,

applicationsdigitalocean
3 Jun 2026
Research

The Unsampled Truth: Psychometrics in SLMs Measure Prompt Artifacts, Not Psychological Constructs

DGX agent

arXiv:2606.03357v1 Announce Type: cross Abstract: When prompting SLMs for psychometric assessments, researchers assume the outputs reflect semantic reasoning. We evaluate this premise across 13 open-w

researcharxiv-cs-ai
3 Jun 2026
Industry

The US and other Five Eyes nations warn that China is flooding online job platforms with fake profiles and offers targeting government and military personnel (Greg Miller/Washington Post)

DGX agent

Greg Miller / Washington Post: The US and other Five Eyes nations warn that China is flooding online job platforms with fake profiles and offers targeting government and military personnel — Nations i

industrytechmeme
3 Jun 2026
Industry

The US data center build-out is falling behind schedule; JP Morgan says 60%+ of data center capacity planned for completion in 2027 isn't yet under construction (Katherine Blunt/Wall Street Journal)

DGX agent

Katherine Blunt / Wall Street Journal: The US data center build-out is falling behind schedule; JP Morgan says 60%+ of data center capacity planned for completion in 2027 isn't yet under construction

industrytechmeme
3 Jun 2026
Applications

The Violation Situation Pattern: A Knowledge-Graph Pattern for Compliance Violations

DGX agent

arXiv:2606.03326v1 Announce Type: new Abstract: Compliance pipelines detect violations as transient query results and do not keep the violation itself as a persistent graph object with review state, a

applicationsarxiv-cs-ai
3 Jun 2026
Industry

The West has created an utterly evil state religion where an accusation of “racism” is the gravest offense that can be committed, even worse…

DGX agent

The West has created an utterly evil state religion where an accusation of “racism” is the gravest offense that can be committed, even worse than rape or murder! So if police show up at a crime scene

industryelon-musk--x
3 Jun 2026
← Previous
1…918919920921922…1911
Next →