AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,013 results
Model Releases

Plant, Persist, Trigger: Sleeper Attack on Large Language Model Agents

DGX agent

arXiv:2605.28201v1 Announce Type: new Abstract: Large Language Model (LLM) agents remain vulnerable to safety threats from the external environment, where attackers inject adversarial content into ext

model-releasesarxiv-cs-ai
28 May 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Restoring the Sweet Spot: Pass-Rate Weighted Self-Distillation for LLM Reasoning

DGX agent

arXiv:2605.27765v1 Announce Type: cross Abstract: Self-Distillation Policy Optimization (SDPO) provides dense token-level credit assignment for reinforcement learning with large language models by lev

safetyarxiv-cs-ai
28 May 2026
Research

ReverseMath: Answer Inversion for Scalable and Verifiable Mathematical Problem Generation

DGX agent

arXiv:2605.27709v1 Announce Type: new Abstract: Mathematical reasoning benchmarks are vital for evaluating large language models (LLMs), but many are static and repeatedly exposed through public evalu

researcharxiv-cs-cl
28 May 2026
Applications

Revisiting Change Detection Methods for their Application to Serac Fall Time-Lapse Monitoring

DGX agent

arXiv:2605.28100v1 Announce Type: cross Abstract: In an era where climate change aggravates environmental uncertainties, the identification and detection of event precursors are becoming crucial to mi

applicationsarxiv-cs-ai
28 May 2026
Research

Robo-Blocks: Generative Scaffolding in End-User Design and Programming of Social Robots

DGX agent

arXiv:2605.28154v1 Announce Type: cross Abstract: Programming social robots is challenging for novice robot programmers due to required expertise in planning, interaction design, and programming. Whil

researcharxiv-cs-ro
28 May 2026
Model Releases

SHIPPED. Mistral Vibe is now the AI agent for long-horizon productivity and coding, and the home for Work mode, Code mode, the CLI, and a br…

DGX agent

Mistral AI has released Mistral Vibe, an AI agent designed for long-horizon productivity and coding tasks, featuring Work mode, Code mode, a CLI, and additional capabilities. The product consolidates

model-releasesmistral-ai--x
28 May 2026
Agents

slide from

DGX agent

slide from The Redpoint InfraRed 100 is now live. These are the companies building the infrastructure that powers everything happening in AI right now, from world models and agent runtimes to the sand

agentsswyx--x
28 May 2026
Research

STFlow: Data-Coupled Flow Matching for Geometric Trajectory Simulation

DGX agent

arXiv:2505.18647v3 Announce Type: replace-cross Abstract: Simulating trajectories of dynamical systems is a fundamental problem in a wide range of fields such as molecular dynamics, biochemistry, and

researcharxiv-cs-ai
28 May 2026
Model Releases

The Abstraction Gap in Vision-Language Causal Reasoning

DGX agent

arXiv:2605.28779v1 Announce Type: new Abstract: Vision-language models (VLMs) generate fluent causal explanations, but current evaluations cannot distinguish linguistic plausibility from faithful caus

model-releasesarxiv-cs-cl
28 May 2026
Hardware

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and…

DGX agent

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and fast & lightweight server-side use-cases. > 8B MoE, 1.5B ac

hardwareclem-delangue--x
28 May 2026
Model Releases

Towards Faithful Agentic XAI: A Verification Method and an Open-World Benchmark for Better Model Faithfulness

DGX agent

arXiv:2605.27879v1 Announce Type: new Abstract: Explainable AI (XAI) helps users interpret model behavior and identify potential faults. Agentic XAI systems use Large Language Models (LLMs) to make ex

model-releasesarxiv-cs-ai
28 May 2026
Research

Unifying Low Dimensional Spectra in Deep Learning

DGX agent

arXiv:2404.06106v3 Announce Type: replace Abstract: Low dimensional structures appear ubiquitously in the eigenspectra of deep learning matrices in classification networks trained in the overparameter

researcharxiv-cs-lg
28 May 2026
Model Releases

VeriTrip: A Verifiable Benchmark for Travel Planning Agents over Unstructured Web Corpora

DGX agent

arXiv:2605.28683v1 Announce Type: new Abstract: Existing benchmarks have laid the foundation for travel planning agents by establishing API-centric paradigms. However, as the capabilities of Autonomou

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

When Interpretability Is Unequally Distributed: Fairness in Hybrid Interpretable Models

DGX agent

arXiv:2605.28626v1 Announce Type: new Abstract: Hybrid interpretable models combine a transparent component with a black-box model by assigning some examples to the former and deferring the rest to th

model-releasesarxiv-cs-lg
28 May 2026
Agents

Writer helps solve brand consistency for enterprise marketing at scale

DGX agent

Writer Inc., an enterprise artificial intelligence agent platform used by leading enterprise brands to deliver their voice, today announced new infrastructure aimed at enforcing style, terminology and

agentssiliconangle
28 May 2026
Research

Assessing Per-Sample Membership Inference Vulnerability without Retraining

DGX agent

arXiv:2602.15919v2 Announce Type: replace-cross Abstract: Recent work in the privacy literature shows that sample-targeted membership inference attacks (MIAs) significantly outperform untargeted appro

researcharxiv-cs-ai
27 May 2026
Applications

Been using Grok Build these past few days, and the thing that really got me hooked is Imagine and Imagine Video. I built a full dinosaur enc…

DGX agent

Been using Grok Build these past few days, and the thing that really got me hooked is Imagine and Imagine Video. I built a full dinosaur encyclopedia site — every image, every video clip on it, all ge

applicationselon-musk--x
27 May 2026
Safety

Beyond the Data Mesh Illusion: Designing Modern AI-augmented Lakehouses to Bridge the Gap Between Theory and Practice

DGX agent

arXiv:2605.27131v1 Announce Type: cross Abstract: Enterprise data platforms face an enduring tension between domain self-service and holistic governance. The data mesh paradigm proposed decentralized

safetyarxiv-cs-ai
27 May 2026
Model Releases

Black-box Membership Inference Attacks on the Pre-training Data of Image-generation Models

DGX agent

arXiv:2605.27020v1 Announce Type: cross Abstract: The rapid advancement of diffusion-based image generation models has raised serious concerns regarding potential copyright and privacy infringements i

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

CleanSurvival: Automated data preprocessing for time-to-event models using reinforcement learning

DGX agent

arXiv:2502.03946v5 Announce Type: replace Abstract: Data preprocessing is often paid little attention in machine learning, despite its potentially significant impact on model performance. While automa

model-releasesarxiv-cs-lg
27 May 2026
Safety

Completely agree, @hwchase17 ! This is the meta-level breakthrough we've all been waiting for. Self-optimizing loops finally feel production…

DGX agent

Completely agree, @hwchase17 ! This is the meta-level breakthrough we've all been waiting for. Self-optimizing loops finally feel production-ready because LangSmith Engine turns evaluation from a manu

safetyharrison-chase--x
27 May 2026
Safety

Dimensional Distribution Emotion State: Leveraging Valence and Arousal as a Common Embedding Space for Visual Emotion Analysis

DGX agent

arXiv:2605.26262v1 Announce Type: new Abstract: Museums are important sites for the dissemination of culture and art. They are institutions rooted in history and tradition; their exhibitions are often

safetyarxiv-cs-cv
27 May 2026
Model Releases

EdgeFlow: Edge-Map Augmented VLM-Based Flowchart Processing for Industrial Requirements Engineering

DGX agent

arXiv:2605.27332v1 Announce Type: cross Abstract: Flowcharts are widely used in industrial requirements, but usually remain embedded as static images. Vision Language Models (VLMs) show promise in the

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

ENPMR-Bench: Benchmarking Proactive Memory Retrieval for Emotional Support Agents

DGX agent

arXiv:2605.27240v1 Announce Type: new Abstract: Memory-augmented language agents are increasingly deployed in affective applications such as emotional support, where understanding and responding to us

model-releasesarxiv-cs-cl
27 May 2026
Safety

Few-shot Cross-country Generalization of Tabular Machine Learning and Foundation Models for Childhood Anemia Prediction under Distribution Shift

DGX agent

arXiv:2605.26589v1 Announce Type: cross Abstract: Childhood anemia affects around 40% of children aged 6-59 months globally and arises from heterogeneous factors, limiting model generalizability. We e

safetyarxiv-cs-ai
27 May 2026
Model Releases

FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies

DGX agent

arXiv:2605.27284v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are increasingly expected to not only complete robot tasks, but also follow human instructions about how those tas

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

From PDF to RAG-Ready: Evaluating Document Conversion Frameworks for Domain-Specific Question Answering

DGX agent

arXiv:2604.04948v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems depend critically on the quality of document preprocessing, yet no prior study has evaluated PDF

model-releasesarxiv-cs-ai
27 May 2026
Research

Granuscore: A Reference-Free Measure of Granularity for Text Analysis and Question Answering

DGX agent

arXiv:2605.26620v1 Announce Type: new Abstract: Natural language conveys information at varying levels of granularity, from fine-grained references to broad descriptions. While granularity is fundamen

researcharxiv-cs-cl
27 May 2026
Agents

https://hermes-agent.nousresearch.com/docs/user-guide/features/mcp#catalog-one-click-install-for-nous-approved-mcps

DGX agent

Nous Research introduced a one-click installation feature in Hermes Agent that allows users to easily install Model Context Protocol (MCP) servers from a curated catalog of Nous-approved integrations.

agentsnous-research--x
27 May 2026
Model Releases

If we had done everything I suggested in my 2020 arXiv article “The Next Decade in AI”, we might actually have reached AGI by now. In the la…

DGX agent

If we had done everything I suggested in my 2020 arXiv article “The Next Decade in AI”, we might actually have reached AGI by now. In the last three years, after a detour driven by the false promise o

model-releasesgary-marcus--x
27 May 2026
Applications

Implementation of Big Data Analytics for Diabetes Management: Needs Assessment in the Rwanda Healthcare System

DGX agent

arXiv:2605.26786v1 Announce Type: cross Abstract: Diabetes is a chronic metabolic disease that can lead to serious health problems if not diagnosed and managed early. Big Data Analytics (BDA) and mach

applicationsarxiv-cs-ai
27 May 2026
Model Releases

InterSketch: An Interleaved Reasoning Model with Self-correcting Visual Sketch and Stepwise Reward

DGX agent

arXiv:2605.26520v1 Announce Type: cross Abstract: While vision-language models (VLMs) have exhibited multi-turn visual reasoning capabilities, their reasoning trajectories remain relatively shallow an

model-releasesarxiv-cs-ai
27 May 2026
Safety

Intuitions of Machine Learning Researchers about Transfer Learning for Medical Image Classification

DGX agent

arXiv:2510.00902v2 Announce Type: replace Abstract: Transfer learning is crucial for medical imaging, yet the selection of source datasets often relies on researchers' intuition rather than systematic

safetyarxiv-cs-cv
27 May 2026
Safety

LearnedCache: An eBPF-Integrated Perceptron-Based Eviction Policy for the Linux Page Cache

DGX agent

arXiv:2605.26168v1 Announce Type: cross Abstract: Linux is the foundation of the digital age, accounting for the majority of the cloud and mobile OS markets. Any device that runs Linux uses the Linux

safetyarxiv-cs-lg
27 May 2026
Research

Lost in Sampling: Assessing Lexical Reachability in LLMs via the Word Coverage Score (WCS)

DGX agent

arXiv:2605.27268v1 Announce Type: cross Abstract: Modern Large Language Models (LLMs) are often criticized for producing repetitive and homogeneous text, despite possessing vast latent vocabularies. W

researcharxiv-cs-ai
27 May 2026
Model Releases

Maat: The Agentic Legal Research Assistant for Competition Protection

DGX agent

arXiv:2605.27331v1 Announce Type: new Abstract: Competition law experts conducting legal research must review extensive volumes of cases, decisions, and judicial reports to identify precedents and ass

model-releasesarxiv-cs-ai
27 May 2026
Applications

Managed deep agents is the easiest way to build and deploy long horizon agents Private preview, dm me if you want access

DGX agent

Managed deep agents is the easiest way to build and deploy long horizon agents Private preview, dm me if you want access Managed Deep Agents is built for agents that need to work over long time horizo

applicationsharrison-chase--x
27 May 2026
Research

Managing Uncertainty in LLM-Generated Procedural Knowledge for Virtual Laboratory Planning

DGX agent

arXiv:2605.26333v1 Announce Type: new Abstract: Educational virtual laboratories can make experimental training more scala-ble, adaptive, and accessible, especially when students have limited access t

researcharxiv-cs-ai
27 May 2026
Tutorials

MatFormBench: A Benchmarking Evaluation Framework for Target-Driven Materials Formulation

DGX agent

arXiv:2605.26741v1 Announce Type: cross Abstract: Inverse design of materials has significantly advanced target-driven formulation optimization, yet existing materials machine learning benchmarks rema

tutorialsarxiv-cs-ai
27 May 2026
Model Releases

On the Hidden Costs of Counterfactual Knowledge Training in LLM Unlearning

DGX agent

arXiv:2605.27083v1 Announce Type: new Abstract: Counterfactual tuning (CFT) has emerged as a promising paradigm for Large Language Model (LLM) unlearning by training models to generate alternative fic

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis

DGX agent

arXiv:2605.27258v1 Announce Type: cross Abstract: Building state-of-the-art text-to-speech (TTS) systems typically demands millions of hours of proprietary data and complex multi-stage architectures,

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

PitchBench: Measuring Pitch Hearing in Audio-Language Models

DGX agent

arXiv:2605.26176v1 Announce Type: cross Abstract: Audio-language models (ALMs) are increasingly used in real-world applications that require understanding music, from music tutoring and transcription

model-releasesarxiv-cs-ai
27 May 2026
Agents

PolyFusionAgent: A Multimodal Foundation Model and Autonomous AI Assistant for Polymer Property Prediction and Inverse Design

DGX agent

arXiv:2605.26543v1 Announce Type: new Abstract: Polymer discovery is central to fields ranging from energy storage to biomedicine, but it is hindered by an astronomically large chemical design space a

agentsarxiv-cs-ai
27 May 2026
Safety

ReasonOps: A Unified Operational Paradigm for Trustworthy Verified LLM Reasoning

DGX agent

arXiv:2605.27014v1 Announce Type: cross Abstract: Large Language Models (LLMs) have transformed artificial intelligence from primarily generative systems into increasingly capable reasoning agents. Re

safetyarxiv-cs-ai
27 May 2026
Research

RepoMirage: Probing Repository Context Reasoning in Code Agents with Perturbations

DGX agent

arXiv:2605.26177v1 Announce Type: cross Abstract: Code agents are currently having skillful performance on repository-level software engineering benchmarks, but it remains unclear whether success on e

researcharxiv-cs-ai
27 May 2026
Local Ai

REVERSE: Reinforcing Evidence Verification and Search for Agentic Image geo-localization

DGX agent

arXiv:2605.26861v1 Announce Type: new Abstract: Image geo-localization aims to determine where a photograph was taken, a task that often requires more than recognizing visible landmarks. Human experts

local-aiarxiv-cs-cv
27 May 2026
Model Releases

RoadGIE: Towards A Global-Scale Aerial Benchmark for Generalizable Interactive Road Extraction

DGX agent

arXiv:2605.26862v1 Announce Type: new Abstract: Accurate road segmentation from aerial imagery is fundamental to many geospatial applications. However, existing datasets often suffer from limited scen

model-releasesarxiv-cs-cv
27 May 2026
Applications

Semantic Gradients Interactions in SSD: A Case Study in Racial Identity and Hate Speech

DGX agent

arXiv:2605.27322v1 Announce Type: new Abstract: We introduce interaction SSD, an extension of Supervised Semantic Differential that models how semantic meaning varies across moderators such as groups,

applicationsarxiv-cs-cl
27 May 2026
← Previous
1…167168169170171…209
Next →