AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,005 results
12 May 2026

SmartEval: A Benchmark for Evaluating LLM-Generated Smart Contracts from Natural Language Specifications

Model ReleasesDGX agent

arXiv:2605.09610v1 Announce Type: cross Abstract: We introduce SmartEval, a benchmark for systematically evaluating the quality of Solidity smart contracts generated by large language models (LLMs) fr

Source: Anthropic is in advanced talks to acquire New York-based Stainless, which helps developers generate SDKs from APIs, for at least $300M (The Information)

IndustryDGX agent

The Information: Source: Anthropic is in advanced talks to acquire New York-based Stainless, which helps developers generate SDKs from APIs, for at least 300M — Anthropic is in advanced talks to acqui

Sources: Wispr Flow developer Wispr AI is in talks to raise a round that could more than double its valuation to 2B; source: the round is set to total ~260M (Bloomberg)

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
IndustryDGX agent

Bloomberg: Sources: Wispr Flow developer Wispr AI is in talks to raise a round that could more than double its valuation to 2B; source: the round is set to total ~260M — Wispr AI Inc., the developer b

Spherical Boltzmann machines: a solvable theory of learning and generation in energy-based models

Model ReleasesDGX agent

arXiv:2605.09031v1 Announce Type: new Abstract: Energy-based models (EBMs) are flexible generative architectures inspired by statistical physics, but their learning and generative properties remain po

Steerable but Not Decodable: Function Vectors Operate Beyond the Logit Lens

Model ReleasesDGX agent

arXiv:2604.02608v2 Announce Type: replace Abstract: Activation steering presupposes that task-relevant behaviors correspond to linear directions in activation space -- directions that should both stee

Strategic Exploitation in LLM Agent Markets: A Simulation Framework for E-Commerce Trust

Model ReleasesDGX agent

arXiv:2605.10059v1 Announce Type: new Abstract: Agent-based modeling (ABM) has long been used in economics to study human behavior, and large language model (LLM) agents now enable new forms of social

Survey on Disaster Management Datasets for Remote Sensing Based Emergency Applications

ResearchDGX agent

arXiv:2605.08196v1 Announce Type: new Abstract: Recent natural disasters have highlighted the urgent need for efficient data-driven approaches to disaster management. Machine learning (ML) and deep le

The 9 biggest new features in Android 17

IndustryDGX agent

Would it shock you to hear that Android 17 is filled with new AI-enabled features, like improved dictation and vibe-coded widgets? Fortunately, that's not all. The platform is getting non-AI updates t

The Agent Use of Agent Beings: Agent Cybernetics Is the Missing Science of Foundation Agents

AgentsDGX agent

arXiv:2605.10754v1 Announce Type: new Abstract: LLM-based foundation agents that perceive, reason, and act across thousands of reasoning steps are rapidly becoming the dominant paradigm for deploying

The Association of Transformer-based Sentiment Analysis with Symptom Distress and Deterioration in Routine Psychotherapy Care

ResearchDGX agent

arXiv:2605.09838v1 Announce Type: new Abstract: Sentiment analysis has been of long-standing interest in psychotherapy research. Recently, the Transformer deep learning architecture has produced text-

the Mini Shai-Hulud attack is scary because it attacks new AI coding workflows like CI, editor hooks, agent configs, etc

AgentsDGX agent

The Mini Shai-Hulud attack targets emerging AI-assisted development workflows by compromising multiple integration points including continuous integration systems, code editor hooks, and AI agent conf

The Pokemon Theorem and other Fairness Impossibility Results

SafetyDGX agent

arXiv:2605.09221v1 Announce Type: cross Abstract: Fairness impossibility results often look like distinct scalar incompatibility statements. We show that several share one RKHS geometry: fairness crit

There will be no AI jobpocalypse. The story that AI will lead to massive unemployment is stoking unnecessary fear. AI — like any other techn…

SafetyDGX agent

There will be no AI jobpocalypse. The story that AI will lead to massive unemployment is stoking unnecessary fear. AI — like any other technology — does affect jobs, but telling overblown stories of l

thoughts after doing a bunch of synthetic data gen for eval + environment building - LLMs are incredible projections of the world bundled in…

AgentsDGX agent

thoughts after doing a bunch of synthetic data gen for eval + environment building - LLMs are incredible projections of the world bundled into a set of weights - but doing targeted extraction of certa

Topological Data Analysis Applications in Natural Language Processing: A Survey

ApplicationsDGX agent

arXiv:2411.10298v5 Announce Type: replace Abstract: The surge of data available on the Internet has driven the adoption of a wide range of computational methods for analyzing and extracting insights f

Towards a Certificate of Trust: Task-Aware OOD Detection for Scientific AI

ResearchDGX agent

arXiv:2509.25080v3 Announce Type: replace Abstract: Data-driven models are increasingly adopted in critical scientific fields like weather forecasting and fluid dynamics. These methods can fail on out

Towards a Virtual Neuroscientist: Autonomous Neuroimaging Analysis via Multi-Agent Collaboration

AgentsDGX agent

arXiv:2605.09366v1 Announce Type: new Abstract: Transforming neuroimaging data into clinically actionable biomarkers is a knowledge-intensive and labor-intensive process. Standardized workflows such a

Trajectory-Consistent Flow Matching for Robust Visuomotor Policy Learning

SafetyDGX agent

arXiv:2605.08511v1 Announce Type: new Abstract: Flow matching policies learn continuous velocity fields that transport noise to actions, enabling fast deterministic inference for robot manipulation. H

TrajPrism: A Multi-Task Benchmark for Language-Grounded Urban Trajectory Understanding

Model ReleasesDGX agent

arXiv:2605.10782v1 Announce Type: new Abstract: Urban mobility is naturally expressed both as trajectories in space and as natural-language descriptions of travel intent, constraints, and preferences.

Transformer autoencoder with local attention for sparse and irregular time series with application on risk estimation

ApplicationsDGX agent

arXiv:2605.08914v1 Announce Type: cross Abstract: This paper introduces a framework specifically designed for sparse and irregular time series {risk estimation}. It is based on a Transformer Autoencod

Transforming the Use of Earth Observation Data: Exascale Training of a Generative Compression Model with Historical Priors for up to 10,000x Data Reduction

ResearchDGX agent

arXiv:2605.08633v1 Announce Type: cross Abstract: Earth observation is becoming one of the largest data-producing activities in science, yet current pipelines still treat compression as a storage and

Trustworthy AI: Ensuring Reliability and Accountability from Models to Agents

SafetyDGX agent

arXiv:2605.08964v1 Announce Type: new Abstract: In this thesis, we develop algorithms with theoretical guarantees for ensuring reliability and accountability of Machine Learning (ML) systems. As ML sy

Universal Feature Selection with Noisy Observations and Weak Symmetry Conditions

ResearchDGX agent

arXiv:2605.09396v1 Announce Type: cross Abstract: This paper relaxes the restrictive symmetry conditions adopted in [4], [5] and extends their universal feature selection framework to accommodate nois

V-ABS: Action-Observer Driven Beam Search for Dynamic Visual Reasoning

SafetyDGX agent

arXiv:2605.10172v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have achieved remarkable success in general perception, yet complex multi-step visual reasoning remains a per

Voice choice shapes how an agent feels to users, from fintech support to healthcare intake to entertainment. Try voice finder: https://findt…

AgentsDGX agent

Voice selection significantly impacts user perception and experience across various applications, including financial services support, healthcare intake processes, and entertainment platforms. Togeth

VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning

Model ReleasesDGX agent

arXiv:2605.08146v1 Announce Type: cross Abstract: Multi-model learning has attracted great attention in visual-text tasks. However, visual-tabular data, which plays a pivotal role in high-stakes domai

What Software Engineering Looks Like to AI Agents? -- An Empirical Study of AI-Only Technical Discourse on MoltBook

AgentsDGX agent

arXiv:2605.08380v1 Announce Type: cross Abstract: AI agents are increasingly framed as software-engineering teammates, yet most research studies them inside human-centered workflows. Little is known a

When Agents Overtrust Environmental Evidence: An Extensible Agentic Framework for Benchmarking Evidence-Grounding Defects in LLM Agents

SafetyDGX agent

arXiv:2605.08828v1 Announce Type: new Abstract: Large language model agents increasingly operate through environment-facing scaffolds that expose files, web pages, APIs, and logs. These observations i

When Child Inherits: Modeling and Exploiting Subagent Spawn in Multi-Agent Networks

AgentsDGX agent

arXiv:2605.08460v1 Announce Type: cross Abstract: Since the official release of ChatGPT in 2022, large language models (LLMs) have rapidly evolved from chatbot-style interfaces into agentic systems th

11 May 2026

123D: Unifying Multi-Modal Autonomous Driving Data at Scale

AgentsDGX agent

arXiv:2605.08084v1 Announce Type: cross Abstract: The pursuit of autonomous driving has produced one of the richest sensor data collections in all of robotics. However, its scale and diversity remain

A Reproducible Optimisation Protocol for Calibrating Prompt-Based Large Language Model Workflows in Evidence Synthesis

Model ReleasesDGX agent

arXiv:2605.06937v1 Announce Type: new Abstract: This methods article presents a reproducible calibration workflow for prompt-based large language models (LLMs) in structured evidence-synthesis tasks.

A Systematic Investigation of The RL-Jailbreaker in LLMs

SafetyDGX agent

arXiv:2605.07032v1 Announce Type: cross Abstract: The evolution of generative models from next-token predictors to autonomous engines of complex systems necessitates rigorous safety hardening. Adversa

Activation Differences Reveal Backdoors: A Comparison of SAE Architectures

SafetyDGX agent

arXiv:2605.07324v1 Announce Type: cross Abstract: Backdoor attacks on language models pose a significant threat to AI safety, where models behave normally on most inputs but exhibit harmful behavior w

AffineLens: Capturing the Continuous Piecewise Affine Functions of Neural Networks

ResearchDGX agent

arXiv:2605.06218v2 Announce Type: replace Abstract: Piecewise affine neural networks (PANNs) provide a principled geometric perspective on neural network expressivity by characterizing the input--outp

Agent engineering is hard Deep Agents hides a lot of the weird systems complexity, but still gives you more room to customize the harness th…

AgentsDGX agent

Agent engineering is hard Deep Agents hides a lot of the weird systems complexity, but still gives you more room to customize the harness than almost any agent SDK I've used This is harder to build th

Agentic Coding Needs Proactivity, Not Just Autonomy

SafetyDGX agent

arXiv:2605.06717v1 Announce Type: cross Abstract: Coding agents are rapidly changing the landscape of software development, moving from inline completion to autonomous systems that edit repositories,

Asymmetric On-Policy Distillation: Bridging Exploitation and Imitation at the Token Level

SafetyDGX agent

arXiv:2605.06387v2 Announce Type: replace-cross Abstract: On-policy distillation (OPD) trains a student on its own trajectories with token-level teacher feedback and often outperforms off-policy disti

b9110

Local AiDGX agent

The search results show recent llama.cpp releases and general information but don't contain specific details about the b9110 release. Based on the search patterns and similar recent releases documente

Beyond Reasoning: Reinforcement Learning Unlocks Parametric Knowledge in LLMs

ResearchDGX agent

arXiv:2605.07153v1 Announce Type: new Abstract: Reinforcement learning (RL) has achieved remarkable success in LLM reasoning, but whether it can also improve direct recall of parametric knowledge rema

Beyond the Wrapper: Identifying Artifact Reliance in Static Malware Classifiers using TRUSTEE

TutorialsDGX agent

arXiv:2605.07034v1 Announce Type: cross Abstract: Modern cybersecurity relies heavily on static machine-learning-based malware classifiers. However, transformations such as packing and other non-seman

Bounded Fitting for Expressive Description Logics

ResearchDGX agent

arXiv:2605.07452v1 Announce Type: new Abstract: Bounded fitting is an attractive paradigm for learning logical formulas from labeled data examples that offers PAC-style generalization guarantees and c

Bridging the Last Mile of Circuit Design: PostEDA-Bench, a Hierarchical Benchmark for PPA Convergence and DRC Fixing

Model ReleasesDGX agent

arXiv:2605.06936v1 Announce Type: cross Abstract: LLM-based agents are increasingly applied to the 'last mile' of Electronic Design Automation (EDA): repairing residual sign-off Design Rule Check (DRC

CASCADE: Case-Based Continual Adaptation for Large Language Models During Deployment

AgentsDGX agent

arXiv:2605.06702v1 Announce Type: new Abstract: Large language models (LLMs) have become a central foundation of modern artificial intelligence, yet their lifecycle remains constrained by a rigid sepa

Christoffel-DPS: Optimal sensor placement in diffusion posterior sampling for arbitrary distributions

ApplicationsDGX agent

arXiv:2605.06861v1 Announce Type: new Abstract: State estimation is a critical task in scientific, engineering and control applications. Since the reliability of reconstructions depends on the number

Cloud Storage Rapid: Turbocharged object storage for AI and analytics

Model ReleasesDGX agent

At Google Cloud Next ’26 we announced Cloud Storage Rapid, a family of object storage capabilities for data-intensive workloads like AI and analytics. Out of the gate, Cloud Storage Rapid consists of

Conditional generation of antibody sequences with classifier-guided germline-absorbing discrete diffusion

SafetyDGX agent

arXiv:2605.06720v1 Announce Type: cross Abstract: Antibody therapeutics are among the most successful modern medicines, yet computationally designing antibodies with desirable binding and developabili

Contact-Grounded Policy: Dexterous Visuotactile Policy with Generative Contact Grounding

SafetyDGX agent

arXiv:2603.05687v3 Announce Type: replace Abstract: Contact-rich dexterous manipulation with multi-finger hands remains an open challenge in robotics because task success depends on multi-point contac

Crushed the 4M weekly download mark last week for „@langchain/core“ 💪 let’s go!! 🚀 @LangChain_JS @huntlovell @colifran_

AgentsDGX agent

LangChain's JavaScript core package (@langchain/core) reached 4 million weekly downloads, marking a significant milestone for the JavaScript ecosystem of the LangChain framework. This achievement refl

Discovering Multiagent Learning Algorithms with Large Language Models

SafetyDGX agent

arXiv:2602.16928v3 Announce Type: replace-cross Abstract: Much of the advancement in Multi-Agent Reinforcement Learning (MARL) for imperfect-information games has historically depended on the manual,

Do Joint Audio-Video Generation Models Understand Physics?

Model ReleasesDGX agent

arXiv:2605.07061v1 Announce Type: cross Abstract: Joint audio-video generation models are rapidly approaching professional production quality, raising a central question: do they understand audio-visu

Dooly: Configuration-Agnostic, Redundancy-Aware Profiling for LLM Inference Simulation

HardwareDGX agent

arXiv:2605.07985v1 Announce Type: cross Abstract: Selecting the optimal LLM inference configuration requires evaluation across hardware, serving engines, attention backends, and model architectures, s

DRIP-R: A Benchmark for Decision-Making and Reasoning Under Real-World Policy Ambiguity in the Retail Domain

Model ReleasesDGX agent

arXiv:2605.07699v1 Announce Type: cross Abstract: LLM-based agents are increasingly deployed for routine but consequential tasks in real-world domains, where their behavior is governed by inherently a

Emergence of Distortions in High-Dimensional Guided Diffusion Models

ApplicationsDGX agent

arXiv:2602.00716v4 Announce Type: replace-cross Abstract: Classifier-free guidance (CFG) is the de facto standard for conditional sampling in diffusion models, yet it often reduces sample diversity. U

Exact Is Easier: Credit Assignment for Cooperative LLM Agents

Model ReleasesDGX agent

arXiv:2603.06859v2 Announce Type: replace-cross Abstract: Removing an agent from a cooperative team to measure its contribution seems natural, yet in multi-agent LLM systems this evaluation distorts t

Factored Classifier-Free Guidance

ResearchDGX agent

arXiv:2506.14399v5 Announce Type: replace-cross Abstract: Counterfactual generation aims to simulate realistic hypothetical outcomes under causal interventions. Diffusion models have emerged as a powe

Fine-tuning on your proprietary data is the highest leverage thing you can do. Prompts get copied overnight. A model trained on your data, y…

Model ReleasesDGX agent

Fine-tuning on your proprietary data is the highest leverage thing you can do. Prompts get copied overnight. A model trained on your data, your evals, your edge cases is a strong moat. OpenAI is windi

FlightSense: An End-to-End MLOps Platform for Real-Time Flight Delay Prediction via Rotation-Chain Propagation Features and Agentic Conversational AI

AgentsDGX agent

arXiv:2605.07364v1 Announce Type: new Abstract: Flight delays impose cascading operational and financial burdens across the aviation network, costing the U.S. economy billions of dollars annually by d

From Canopy to Collision: A Hybrid Predictive Framework for Identifying Risk Factors in Tree-Involved Traffic Crashes

ResearchDGX agent

arXiv:2605.06684v1 Announce Type: new Abstract: Tree-involved crashes represent a critical subset of run-off-road (ROR) collisions, often resulting in fatal or severe injuries due to high-energy impac

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems

ResearchDGX agent

arXiv:2506.04565v2 Announce Type: replace-cross Abstract: Compound AI Systems (CAIS) are an emerging paradigm that integrates large language models (LLMs) with external components, including retriever

Future-proof your data strategy: AlloyDB adds PostgreSQL 18 and new Extended Support

SafetyDGX agent

As you look out at your 2026 infrastructure roadmap, your goal is to balance the need for rapid innovation with operational stability. You shouldn't have to choose between adopting the latest database

← Previous
1…144145146147148…167
Next →