AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,013 results
28 May 2026

EchoAvatar: Real-time Generative Avatar Animation from Audio Streams

ResearchDGX agent

arXiv:2605.28272v1 Announce Type: new Abstract: Real-time synthesis of high-fidelity 3D character motion from audio is a pivotal component for next-generation interactive avatars and virtual assistant

Eliot: Interactively nderline{E}xploring Fast-Changing Scientific nderline{Li}terature Trends with nderline{O}nline Danderline{t}a and Learning

ResearchDGX agent

arXiv:2605.27610v1 Announce Type: cross Abstract: The rapid growth of scientific publishing has made it increasingly difficult to track how fast-moving areas evolve. Search engines and LLM-based assis

Evals shape agent behavior. Every eval is a vector that shifts the behavior of your agentic system. More evals ≠ better agents. Instead, bui…

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Evals shape agent behavior. Every eval is a vector that shifts the behavior of your agentic system. More evals ≠ better agents. Instead, build targeted evals that reflect desired behaviors in producti

Evolving Dataflow to process massive datasets for machine learning

Model ReleasesDGX agent

Google created MapReduce more than 20 years ago to solve the scaling problems in data processing that the then young company was running into. The AI era that we are in now demands efficient, large-sc

Extrapolative Weight Averaging Reveals Correctness-Efficiency Frontiers in Code RL

AgentsDGX agent

arXiv:2605.28751v1 Announce Type: cross Abstract: Linear interpolation between fine-tuned checkpoints has been shown to trace the Pareto front between competing objectives, but whether extrapolative w

FBI says Google engineer used internal search data to win $1.2M on Polymarket

IndustryDGX agent

Federal prosecutors charged a Google software engineer with making roughly $1.2 million in profits from bets on the prediction market platform Polymarket by using confidential insider information he l

ForestHG-Trace: Traceable Long-Horizon Ecological Reasoning over Large-Scale Forest Scenes

Model ReleasesDGX agent

arXiv:2605.27590v1 Announce Type: new Abstract: Remote sensing question answering (RS-QA) often requires more than direct semantic prediction, especially in large-scale forest scenes where ecological

From Knowing to Doing: A Memory-Controlled Benchmark for LLM Trading Agents on Stock Markets

Model ReleasesDGX agent

arXiv:2605.28359v1 Announce Type: new Abstract: Evaluating whether large language model (LLM) agents can profit in capital markets is increasingly framed as end-to-end trading: place an agent in a his

Good value for money

IndustryDGX agent

Good value for money We gave Grok Build 0.1 one prompt: build a webhook delivery service in TypeScript, Bun, and SQLite. It planned it, built it, and shipped a working demo. Total cost: $1.65. Zero to

How Endava builds an agentic organization with Codex

AgentsDGX agent

Endava, a software services company, leverages OpenAI's Codex to transform its organizational operations into an agentic model where AI agents autonomously handle tasks and decision-making. The approa

I have spoken to lots of companies that blew through their token budget for the year in months, but a half a billion dollars on internal emp…

ApplicationsDGX agent

Ethan Mollick discusses companies that rapidly depleted their annual token budgets within months, with some spending hundreds of millions of dollars on internal employee use of large language models.

I'm proud to share that @Glean has surpassed 300M ARR, just five months after crossing 200M and growing ~3x over the past 15 months. This …

Model ReleasesDGX agent

I'm proud to share that @Glean has surpassed 300M ARR, just five months after crossing 200M and growing ~3x over the past 15 months. This is an exciting milestone for Glean, and it's a signal about wh

Insurance Pricing Optimization via Off-Policy Evaluation

SafetyDGX agent

arXiv:2605.28327v1 Announce Type: cross Abstract: Traditional insurance pricing relies on risk-based principles that ensure actuarial fairness and solvency but do not explicitly account for policyhold

Interpretability-Guided Layer Selection over Subspace Projection: SAEs as Stethoscopes, Not Scalpels, for Raw Task Vector Model Editing

Model ReleasesDGX agent

arXiv:2605.28649v1 Announce Type: cross Abstract: LLMs increasingly require surgical model editing to enhance domain-specific capabilities without incurring the computational cost or catastrophic forg

KSAFE-MM: A Multimodal Safety Benchmark via Localized Contextualization for Korean Cultural Risks

Model ReleasesDGX agent

arXiv:2605.28013v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) exacerbate safety risks by introducing vulnerabilities across multiple modalities, such as language and vision.

Large Language Models as Automatic Annotators and Annotation Adjudicators for Fine-Grained Opinion Analysis

AgentsDGX agent

arXiv:2601.16800v3 Announce Type: replace Abstract: Fine-grained opinion analysis of text provides a detailed understanding of expressed sentiments, including the addressed entity. Although this level

Learning Logical Operations for Arbitrary Quantum Error Correction Codes

ResearchDGX agent

arXiv:2605.28162v1 Announce Type: cross Abstract: Logical operations are essential for quantum computation within quantum error-correcting codes. However, discovering their physical realizations is ch

Learning with Importance Weighted Variational Inference

SafetyDGX agent

arXiv:2410.12035v2 Announce Type: replace-cross Abstract: Several variational bounds involving importance weighting ideas generalize the Evidence Lower BOund (ELBO) for marginal likelihood optimizatio

Local Reachy Mini conversations wireless looks like magic! You can bring your friend around the house and get WOW effect from anyone! 🔥 Tha…

Model ReleasesDGX agent

Local Reachy Mini conversations wireless looks like magic! You can bring your friend around the house and get WOW effect from anyone! 🔥 Thanks @andimarafioti for the blog post on how to set this up: h

Mathematical Modelling of Ethical AI Use in Higher Education: A Coordination Game Framework for Future-Facing Learning

SafetyDGX agent

arXiv:2605.27400v1 Announce Type: cross Abstract: The rapid uptake of generative artificial intelligence (AI) in higher education is reshaping assessment practices and intensifying concerns around aca

models being conscious would be harmful for humanity. it would encroach on our status and dignity. it would limit the type of things we can …

ApplicationsDGX agent

models being conscious would be harmful for humanity. it would encroach on our status and dignity. it would limit the type of things we can do with them and use them for. it would vastly accelerate hu

Multi-Agent LLM-based Metamorphic Testing for REST APIs

AgentsDGX agent

arXiv:2605.28321v1 Announce Type: cross Abstract: As REST APIs become an increasingly significant part of software systems, their validation is becoming more critical. Hence, testing and uncovering un

On the Subgaussianity of Quantized Linear Maps: An AI-Assisted Note

Model ReleasesDGX agent

arXiv:2605.27563v1 Announce Type: cross Abstract: This short note presents a dimension-independent subgaussian concentration bound for Gaussian vectors under coordinate-wise nonlinear mappings. Discov

Pittsburgh-based Gray Swan, which stress-tests AI models for top frontier AI labs, raised a 40M Series A at a 200M valuation co-led by Wing VC and Madrona (Rashi Shrivastava/Forbes)

Model ReleasesDGX agent

Rashi Shrivastava / Forbes: Pittsburgh-based Gray Swan, which stress-tests AI models for top frontier AI labs, raised a 40M Series A at a 200M valuation co-led by Wing VC and Madrona — Gray Swan works

Plant, Persist, Trigger: Sleeper Attack on Large Language Model Agents

Model ReleasesDGX agent

arXiv:2605.28201v1 Announce Type: new Abstract: Large Language Model (LLM) agents remain vulnerable to safety threats from the external environment, where attackers inject adversarial content into ext

Restoring the Sweet Spot: Pass-Rate Weighted Self-Distillation for LLM Reasoning

SafetyDGX agent

arXiv:2605.27765v1 Announce Type: cross Abstract: Self-Distillation Policy Optimization (SDPO) provides dense token-level credit assignment for reinforcement learning with large language models by lev

ReverseMath: Answer Inversion for Scalable and Verifiable Mathematical Problem Generation

ResearchDGX agent

arXiv:2605.27709v1 Announce Type: new Abstract: Mathematical reasoning benchmarks are vital for evaluating large language models (LLMs), but many are static and repeatedly exposed through public evalu

Revisiting Change Detection Methods for their Application to Serac Fall Time-Lapse Monitoring

ApplicationsDGX agent

arXiv:2605.28100v1 Announce Type: cross Abstract: In an era where climate change aggravates environmental uncertainties, the identification and detection of event precursors are becoming crucial to mi

Robo-Blocks: Generative Scaffolding in End-User Design and Programming of Social Robots

ResearchDGX agent

arXiv:2605.28154v1 Announce Type: cross Abstract: Programming social robots is challenging for novice robot programmers due to required expertise in planning, interaction design, and programming. Whil

SHIPPED. Mistral Vibe is now the AI agent for long-horizon productivity and coding, and the home for Work mode, Code mode, the CLI, and a br…

Model ReleasesDGX agent

Mistral AI has released Mistral Vibe, an AI agent designed for long-horizon productivity and coding tasks, featuring Work mode, Code mode, a CLI, and additional capabilities. The product consolidates

slide from

AgentsDGX agent

slide from The Redpoint InfraRed 100 is now live. These are the companies building the infrastructure that powers everything happening in AI right now, from world models and agent runtimes to the sand

STFlow: Data-Coupled Flow Matching for Geometric Trajectory Simulation

ResearchDGX agent

arXiv:2505.18647v3 Announce Type: replace-cross Abstract: Simulating trajectories of dynamical systems is a fundamental problem in a wide range of fields such as molecular dynamics, biochemistry, and

The Abstraction Gap in Vision-Language Causal Reasoning

Model ReleasesDGX agent

arXiv:2605.28779v1 Announce Type: new Abstract: Vision-language models (VLMs) generate fluent causal explanations, but current evaluations cannot distinguish linguistic plausibility from faithful caus

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and…

HardwareDGX agent

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and fast & lightweight server-side use-cases. > 8B MoE, 1.5B ac

Towards Faithful Agentic XAI: A Verification Method and an Open-World Benchmark for Better Model Faithfulness

Model ReleasesDGX agent

arXiv:2605.27879v1 Announce Type: new Abstract: Explainable AI (XAI) helps users interpret model behavior and identify potential faults. Agentic XAI systems use Large Language Models (LLMs) to make ex

Unifying Low Dimensional Spectra in Deep Learning

ResearchDGX agent

arXiv:2404.06106v3 Announce Type: replace Abstract: Low dimensional structures appear ubiquitously in the eigenspectra of deep learning matrices in classification networks trained in the overparameter

VeriTrip: A Verifiable Benchmark for Travel Planning Agents over Unstructured Web Corpora

Model ReleasesDGX agent

arXiv:2605.28683v1 Announce Type: new Abstract: Existing benchmarks have laid the foundation for travel planning agents by establishing API-centric paradigms. However, as the capabilities of Autonomou

When Interpretability Is Unequally Distributed: Fairness in Hybrid Interpretable Models

Model ReleasesDGX agent

arXiv:2605.28626v1 Announce Type: new Abstract: Hybrid interpretable models combine a transparent component with a black-box model by assigning some examples to the former and deferring the rest to th

Writer helps solve brand consistency for enterprise marketing at scale

AgentsDGX agent

Writer Inc., an enterprise artificial intelligence agent platform used by leading enterprise brands to deliver their voice, today announced new infrastructure aimed at enforcing style, terminology and

27 May 2026

Assessing Per-Sample Membership Inference Vulnerability without Retraining

ResearchDGX agent

arXiv:2602.15919v2 Announce Type: replace-cross Abstract: Recent work in the privacy literature shows that sample-targeted membership inference attacks (MIAs) significantly outperform untargeted appro

Been using Grok Build these past few days, and the thing that really got me hooked is Imagine and Imagine Video. I built a full dinosaur enc…

ApplicationsDGX agent

Been using Grok Build these past few days, and the thing that really got me hooked is Imagine and Imagine Video. I built a full dinosaur encyclopedia site — every image, every video clip on it, all ge

Beyond the Data Mesh Illusion: Designing Modern AI-augmented Lakehouses to Bridge the Gap Between Theory and Practice

SafetyDGX agent

arXiv:2605.27131v1 Announce Type: cross Abstract: Enterprise data platforms face an enduring tension between domain self-service and holistic governance. The data mesh paradigm proposed decentralized

Black-box Membership Inference Attacks on the Pre-training Data of Image-generation Models

Model ReleasesDGX agent

arXiv:2605.27020v1 Announce Type: cross Abstract: The rapid advancement of diffusion-based image generation models has raised serious concerns regarding potential copyright and privacy infringements i

CleanSurvival: Automated data preprocessing for time-to-event models using reinforcement learning

Model ReleasesDGX agent

arXiv:2502.03946v5 Announce Type: replace Abstract: Data preprocessing is often paid little attention in machine learning, despite its potentially significant impact on model performance. While automa

Completely agree, @hwchase17 ! This is the meta-level breakthrough we've all been waiting for. Self-optimizing loops finally feel production…

SafetyDGX agent

Completely agree, @hwchase17 ! This is the meta-level breakthrough we've all been waiting for. Self-optimizing loops finally feel production-ready because LangSmith Engine turns evaluation from a manu

Dimensional Distribution Emotion State: Leveraging Valence and Arousal as a Common Embedding Space for Visual Emotion Analysis

SafetyDGX agent

arXiv:2605.26262v1 Announce Type: new Abstract: Museums are important sites for the dissemination of culture and art. They are institutions rooted in history and tradition; their exhibitions are often

EdgeFlow: Edge-Map Augmented VLM-Based Flowchart Processing for Industrial Requirements Engineering

Model ReleasesDGX agent

arXiv:2605.27332v1 Announce Type: cross Abstract: Flowcharts are widely used in industrial requirements, but usually remain embedded as static images. Vision Language Models (VLMs) show promise in the

ENPMR-Bench: Benchmarking Proactive Memory Retrieval for Emotional Support Agents

Model ReleasesDGX agent

arXiv:2605.27240v1 Announce Type: new Abstract: Memory-augmented language agents are increasingly deployed in affective applications such as emotional support, where understanding and responding to us

Few-shot Cross-country Generalization of Tabular Machine Learning and Foundation Models for Childhood Anemia Prediction under Distribution Shift

SafetyDGX agent

arXiv:2605.26589v1 Announce Type: cross Abstract: Childhood anemia affects around 40% of children aged 6-59 months globally and arises from heterogeneous factors, limiting model generalizability. We e

FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies

Model ReleasesDGX agent

arXiv:2605.27284v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are increasingly expected to not only complete robot tasks, but also follow human instructions about how those tas

From PDF to RAG-Ready: Evaluating Document Conversion Frameworks for Domain-Specific Question Answering

Model ReleasesDGX agent

arXiv:2604.04948v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems depend critically on the quality of document preprocessing, yet no prior study has evaluated PDF

Granuscore: A Reference-Free Measure of Granularity for Text Analysis and Question Answering

ResearchDGX agent

arXiv:2605.26620v1 Announce Type: new Abstract: Natural language conveys information at varying levels of granularity, from fine-grained references to broad descriptions. While granularity is fundamen

https://hermes-agent.nousresearch.com/docs/user-guide/features/mcp#catalog-one-click-install-for-nous-approved-mcps

AgentsDGX agent

Nous Research introduced a one-click installation feature in Hermes Agent that allows users to easily install Model Context Protocol (MCP) servers from a curated catalog of Nous-approved integrations.

If we had done everything I suggested in my 2020 arXiv article “The Next Decade in AI”, we might actually have reached AGI by now. In the la…

Model ReleasesDGX agent

If we had done everything I suggested in my 2020 arXiv article “The Next Decade in AI”, we might actually have reached AGI by now. In the last three years, after a detour driven by the false promise o

Implementation of Big Data Analytics for Diabetes Management: Needs Assessment in the Rwanda Healthcare System

ApplicationsDGX agent

arXiv:2605.26786v1 Announce Type: cross Abstract: Diabetes is a chronic metabolic disease that can lead to serious health problems if not diagnosed and managed early. Big Data Analytics (BDA) and mach

InterSketch: An Interleaved Reasoning Model with Self-correcting Visual Sketch and Stepwise Reward

Model ReleasesDGX agent

arXiv:2605.26520v1 Announce Type: cross Abstract: While vision-language models (VLMs) have exhibited multi-turn visual reasoning capabilities, their reasoning trajectories remain relatively shallow an

Intuitions of Machine Learning Researchers about Transfer Learning for Medical Image Classification

SafetyDGX agent

arXiv:2510.00902v2 Announce Type: replace Abstract: Transfer learning is crucial for medical imaging, yet the selection of source datasets often relies on researchers' intuition rather than systematic

LearnedCache: An eBPF-Integrated Perceptron-Based Eviction Policy for the Linux Page Cache

SafetyDGX agent

arXiv:2605.26168v1 Announce Type: cross Abstract: Linux is the foundation of the digital age, accounting for the majority of the cloud and mobile OS markets. Any device that runs Linux uses the Linux

Lost in Sampling: Assessing Lexical Reachability in LLMs via the Word Coverage Score (WCS)

ResearchDGX agent

arXiv:2605.27268v1 Announce Type: cross Abstract: Modern Large Language Models (LLMs) are often criticized for producing repetitive and homogeneous text, despite possessing vast latent vocabularies. W

Maat: The Agentic Legal Research Assistant for Competition Protection

Model ReleasesDGX agent

arXiv:2605.27331v1 Announce Type: new Abstract: Competition law experts conducting legal research must review extensive volumes of cases, decisions, and judicial reports to identify precedents and ass

← Previous
1…133134135136137…167
Next →