AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,023 results
9 Jun 2026

Discovering Data Structures: Nearest Neighbor Search and Beyond

Local AiDGX agent

arXiv:2411.03253v2 Announce Type: replace-cross Abstract: We propose a general framework for end-to-end learning of data structures. Our framework adapts to the underlying data distribution and provid

Edge-Constrained UAV Small-Object Detection with P2 Enhancement and Quantum-Inspired Lightweight Structure Search

ResearchDGX agent

arXiv:2606.09081v1 Announce Type: new Abstract: Unmanned aerial vehicle (UAV) object detection requires compact detectors that retain small-object details under onboard computation and memory constrai

Emergence World: A Platform for Evaluating Long-Horizon Multi-Agent Autonomy

Model ReleasesDGX agent

arXiv:2606.08367v1 Announce Type: cross Abstract: Most evaluations of LLM agents look like exams: a discrete task, a clean environment, a score in minutes or hours. We argue that this approach is mism

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Evaluating AI Investment Strategies

SafetyDGX agent

arXiv:2606.08791v1 Announce Type: cross Abstract: We study the problem of auditing a black-box algorithmic decision-maker from observable inputs and outputs alone. Our main result is an exact decompos

Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting

Model ReleasesDGX agent

arXiv:2606.09809v1 Announce Type: new Abstract: AI evaluation results are produced at scale but reported inconsistently across leaderboards, model cards, benchmark papers, and company blogs. The cost

Exposing Hidden Biases in Text-to-Image Models via Automated Prompt Search

SafetyDGX agent

arXiv:2512.08724v3 Announce Type: replace Abstract: Text-to-image (TTI) diffusion models have achieved remarkable visual quality, yet they have been repeatedly shown to exhibit social biases across se

Fable 5 is now available in Claude Code and Cowork Fable is the best model I have used for coding, by a wide margin. It is a big step up, en…

Model ReleasesDGX agent

Fable 5 is now available in Claude Code and Cowork Fable is the best model I have used for coding, by a wide margin. It is a big step up, enabling less prompts and steers, more efficient token use, be

Filigran launches XTM One to automate threat exposure management with AI agents

Model ReleasesDGX agent

French cybersecurity company Filigran SAS today launched XTM One, an artificial intelligence orchestration layer that automates continuous threat exposure management workflows across its platform. XTM

Geometric Analysis of Magnetic Labyrinthine Stripe Evolution via Deep Learning Segmentation

ResearchDGX agent

arXiv:2509.11485v3 Announce Type: replace-cross Abstract: Labyrinthine stripe patterns are common in many physical systems, yet their lack of long-range order makes quantitative characterization chall

How to unlock true ROI in software development – a deep dive into the latest DORA research

TutorialsDGX agent

How do you prove the business value of generative AI to your teams? Technology and finance leaders need to show the clear business value of AI projects to secure ongoing funding. While measuring retur

Knowledge Graphs and Reasoning LLMs for Finding Simple Yet Effective Transcriptomic Perturbation Predictors

ResearchDGX agent

arXiv:2606.08816v1 Announce Type: cross Abstract: Predicting the effect of an unseen gene knockout perturbation on transcriptomic gene expression remains a highly challenging problem for virtual cell

Large Models for Time Series and Spatio-Temporal Data: A Survey and Outlook

ApplicationsDGX agent

arXiv:2310.10196v3 Announce Type: replace-cross Abstract: Temporal data, including time series and spatio-temporal data, are pervasive in real-world applications. Generated in massive volumes by physi

Learning to lead in a hybrid human-AI enterprise

ApplicationsDGX agent

As adoption of AI agents looks set to surge by as much as 300% in the next two years, leadership teams are carefully considering the implications of a hybrid human-AI workforce. Unlike existing enterp

me January 2025 @politico, timing slightly off: we may soon see “the largest cyberattack in history, taking down, at least for a little whil…

SafetyDGX agent

me January 2025 @politico, timing slightly off: we may soon see “the largest cyberattack in history, taking down, at least for a little while, some sizeable piece of the world’s infrastructure… genera

MOLOT System Card: Malicious Operational Logic Observation Transformer

Model ReleasesDGX agent

arXiv:2606.07792v1 Announce Type: cross Abstract: MOLOT (Malicious Operational Logic Observation Transformer) is a static malicious-code detection system designed for SAST setup where package metadata

Motion planning for hundreds of floating robots

AgentsDGX agent

arXiv:2606.09620v1 Announce Type: new Abstract: Planning collision-free motion for large robot fleets is difficult because collision avoidance induces strong inter-agent coupling that grows rapidly wi

OmniGameArena: A Unified UE5 Benchmark for VLM Game Agents with Improvement Dynamics

Model ReleasesDGX agent

arXiv:2606.09826v1 Announce Type: cross Abstract: Vision-language model (VLM) agents are increasingly deployed in interactive game environments. Yet game benchmarks for VLM agents typically report a s

Optical Music Recognition for Real-World Manuscripts with Synthetic Data

ApplicationsDGX agent

arXiv:2606.09479v1 Announce Type: new Abstract: Optical Music Recognition (OMR) has seen major progress in model design, with end-to-end methods now capable of recognising notation at all levels of co

Orange Lab: Lowering Barriers to Data Mining through Embedded Interactive Workflows

TutorialsDGX agent

arXiv:2606.09239v1 Announce Type: new Abstract: While visual programming of data analysis workflows has become an important vehicle for the democratization of data science, such systems remain largely

OrderDP: A Theoretically Guaranteed Lossless Dynamic Data Pruning Framework

SafetyDGX agent

arXiv:2606.08574v1 Announce Type: cross Abstract: Data pruning (DP), as an oft-stated strategy to alleviate heavy training burdens, reduces the volume of training samples according to a well-defined p

PBSD: Privileged Bayesian Self-Distillation for Long-Horizon Credit Assignment

SafetyDGX agent

arXiv:2606.09348v1 Announce Type: new Abstract: Long-horizon agentic tasks pose a fundamental credit assignment challenge for outcome-base reinforcement learning: trajectory-level rewards verify final

PIPE-Cypher: Automatic Enterprise Benchmark Generation for Text-to-Cypher Systems

Model ReleasesDGX agent

arXiv:2606.08481v1 Announce Type: cross Abstract: Enterprise property graphs vary widely in schema structure, internal terminology, domain assumptions, governance constraints, and user interaction pat

PLAGUE: Plug-and-play framework for Lifelong Adaptive Generation of Multi-turn Exploits

Model ReleasesDGX agent

arXiv:2510.17947v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are improving at an exceptional rate. With the advent of agentic workflows, multi-turn dialogue has become the de

POISE: Position-Aware Undetectable Skill Injection on LLM Agents

Model ReleasesDGX agent

arXiv:2606.07943v1 Announce Type: cross Abstract: Agent skills provide a lightweight mechanism for extending general-purpose agents, but their open format exposes them to skill-poisoning attacks. A pr

RadOT-Eval: Auditable Structured-Evidence Transport for Radiology Report Evaluation

ResearchDGX agent

arXiv:2606.08769v1 Announce Type: cross Abstract: Automatic evaluation is critical for high-stakes text generation, where errors often involve omitted findings, hallucinated content, polarity reversal

RAILS: Verification-Native Clearing For Agentic Commerce

AgentsDGX agent

arXiv:2606.08790v1 Announce Type: new Abstract: Autonomous agents negotiate, purchase, deploy code, and move funds, but no neutral mechanism determines whether they met their delegated obligation, who

REFLECT: Intervention-Supported Error Attribution for Silent Failures in LLM Agent Traces

AgentsDGX agent

arXiv:2606.09071v1 Announce Type: new Abstract: Large language model (LLM) agents now solve complex tasks through long plan-and-execution traces, yet the ability to locate errors in a completed traces

Revisiting Articulated Parts Perception in Robot Manipulation

SafetyDGX agent

arXiv:2606.08103v1 Announce Type: cross Abstract: We are surrounded by various objects with movable, articulated parts, e.g., box, handle, door. An accurate and generalizable perception of articulated

ridiculous oversimplification, from a prominent OpenAI employee no less. does this capture your impression of how the two companies have beh…

SafetyDGX agent

ridiculous oversimplification, from a prominent OpenAI employee no less. does this capture your impression of how the two companies have behaved? The OAI / Anthropic values difference is deeply misund

Routine laboratory trajectories encode the onset of organ-level complications in cancer

ApplicationsDGX agent

arXiv:2606.08538v1 Announce Type: new Abstract: Routine laboratory panels drawn during cancer treatment constitute longitudinal physiological recordings of organ function, yet their temporal structure

SAW: Stage-Aware Dynamic Weighting for Multi-Objective Reinforcement Learning in Large Language Models

SafetyDGX agent

arXiv:2606.07705v1 Announce Type: cross Abstract: Although multi-objective reinforcement learning (MORL) is central to aligning large language models with complex human preferences, the prevailing pra

Scaffold Effects on GAIA: A Controlled Comparison

Model ReleasesDGX agent

arXiv:2606.08529v1 Announce Type: new Abstract: Published agent capability scores conflate what a model can do with what its scaffold lets it do, and the magnitude of this elicitation gap is not well

// Self-Harness: Harnesses That Improve Themselves // (bookmark this one) Most of the agent scaffolds we rely on today are built once and re…

AgentsDGX agent

// Self-Harness: Harnesses That Improve Themselves // (bookmark this one) Most of the agent scaffolds we rely on today are built once and remain frozen or mostly unchanged. The harness, like the skill

Sequential statistical inference for Large Language Models: Representation, validity, and monitoring

SafetyDGX agent

arXiv:2606.07624v1 Announce Type: new Abstract: This discussion argues that sequential statistical inference can naturally contribute to LLM trustworthiness. In deployment, LLM systems are queried rep

SIGA: Self-Evolving Coding-Agent Adapters for Scientific Simulation

AgentsDGX agent

arXiv:2606.09774v1 Announce Type: new Abstract: Advanced scientific simulators expose specialized input languages that turn simulation goals into executable configurations, but learning them can cost

SlideCheck: Guiding Self-Supervised Pretraining of Pathology Foundation Models via Dataset Distributions

ResearchDGX agent

arXiv:2606.07590v1 Announce Type: cross Abstract: Pathology foundation models are pretrained on large streams of WSI-derived patches, while supervision during data construction is often slide-level, s

SNR-ST-Mix: Sample-specific Neighborhood Regression Mixup for Augmented Spatial Transcriptomics Imputation with Deep Neural Network

Local AiDGX agent

arXiv:2606.08712v1 Announce Type: cross Abstract: Purpose: Spatial transcriptomics (ST) enables gene expression measurements within the tissue context. However, these measurements are often noisy, low

SOMA: From Surface Observations to Muscle Anatomy

ResearchDGX agent

arXiv:2606.09246v1 Announce Type: new Abstract: With the growing demand for realistic virtual humans, parametric body models have become a cornerstone of modern medicine, sports, and entertainment app

Sovereign AI for all.

Model ReleasesDGX agent

Cohere advocates for democratizing access to sovereign AI systems, enabling organizations and nations to develop and deploy their own AI models independently rather than relying on centralized provide

Syll: Open-Source Personal Automation with Cross-Surface Execution

AgentsDGX agent

arXiv:2606.07594v1 Announce Type: new Abstract: Personal AI agents must increasingly operate across APIs, shells, web surfaces, and desktop GUIs, yet many systems remain tuned to a single interface an

The ACUTE Protocol: Operationalizing Language Model Activations for Better Calibration, Utility, and Trust

SafetyDGX agent

arXiv:2606.07822v1 Announce Type: cross Abstract: As language models improve and become increasingly deployed to solve a variety of tasks, trustworthiness becomes essential. Calibration is a good prox

The CIFAR Synthetic Evidence Corpus for Detecting AI-Generated Evidence

TutorialsDGX agent

arXiv:2606.07916v1 Announce Type: new Abstract: The growing ability of generative models to produce realistic documents poses a direct challenge to evidentiary workflows in the justice system and the

Toward autocorrection of chemical process flowsheets using large language models

SafetyDGX agent

arXiv:2312.02873v2 Announce Type: replace-cross Abstract: The process engineering domain widely uses Process Flow Diagrams (PFDs) and Process and Instrumentation Diagrams (P&IDs) to represent process

Trajectory Geometry of Transformer Representations Across Layers

ResearchDGX agent

arXiv:2606.09287v1 Announce Type: new Abstract: Understanding how transformer representations evolve across layers, not merely what they encode, remains an open problem in mechanistic interpretability

Try out Devin today! http://devin.ai

AgentsDGX agent

Devin is an AI software engineer developed by Cognition AI that automates coding tasks and assists with software development. The post appears to be a promotional announcement encouraging users to try

Vision-Guided Dual-Arm Humanoid Robotic Disassembly of End-of-Life 18650 Lithium-ion Battery Packs

TutorialsDGX agent

arXiv:2606.08152v1 Announce Type: new Abstract: The growing volume of retired lithium-ion battery packs from electric vehicles and portable electronics calls for automated disassembly that is safe, fl

WeaveBench: A Long-Horizon, Real-World Benchmark for Computer-Use Agents with Hybrid Interfaces

Model ReleasesDGX agent

arXiv:2606.09426v1 Announce Type: new Abstract: Computer-use agents (CUAs) increasingly operate in runtimes that combine visual desktop control, command-line execution, code editing, browsers, and ext

Web Agents Should Use Typed Actions Instead of Click-Based Browsing

SafetyDGX agent

arXiv:2602.17245v2 Announce Type: replace Abstract: This position paper argues that building a reliable agentic Web requires shifting from low-level interaction primitives to typed actions supported b

What it feels like to work with Mythos

Model ReleasesDGX agent

This article by Ethan Mollick describes the user experience and practical workflow of working with Mythos, an AI system. It likely covers the system's capabilities, interface, strengths, limitations,

Zscaler launches AI Broker and Endpoint AI Security for agents

Model ReleasesDGX agent

Zscaler Inc. today unveiled a set of products designed to secure autonomous artificial intelligence agents, with the cybersecurity company claiming it has built the industry’s first complete zero-trus

8 Jun 2026

3DMorph: Single-Image-Guided Local 3D Shape Editing and Morphing

Model ReleasesDGX agent

arXiv:2606.07115v1 Announce Type: new Abstract: Despite recent progress in 3D generation, intuitive editing of existing shapes remains limited. Unlike images, which benefit from well-established inpai

A Conformation-Centric Generative Foundation Model for Linear Polymer Modeling and Design

ResearchDGX agent

arXiv:2510.16023v2 Announce Type: replace Abstract: Linear polymers, macromolecules formed from monomers covalently bonded into continuous chains, underpin countless technologies and are indispensable

A French engineer who lives quietly in Paris has spent 30 years writing software that the entire internet now runs on without knowing his na…

Model ReleasesDGX agent

A French engineer who lives quietly in Paris has spent 30 years writing software that the entire internet now runs on without knowing his name. He wrote the code that streams every YouTube video, ever

A machine-learning-assisted progressive digit-randomness screening framework for detecting non-random patterns in raw numerical research data

ApplicationsDGX agent

arXiv:2606.07128v1 Announce Type: new Abstract: Raw numerical datasets remain less systematically examined in integrity screening than images, plagiarism, or summary-statistic inconsistencies. We deve

A Multi-Operator Mixed-Reality Interface for Multi-Robot Control and Coordination: Co-Located and Private Workspace Collaboration

ResearchDGX agent

arXiv:2606.07013v1 Announce Type: new Abstract: Multi-operator control of robot teams requires not only access to the same mission information, but also mechanisms for maintaining shared awareness and

Agentic Large Language Models for Automated Structural Analysis of 3D Frame Systems

AgentsDGX agent

arXiv:2606.06525v1 Announce Type: cross Abstract: Large language models (LLMs) have emerged as powerful foundation models with strong reasoning capabilities across domains. Beyond reactive text genera

AI-Driven Test Case Generation from Natural Language Requirements: A Survey of Techniques and Research Gaps

ResearchDGX agent

arXiv:2606.06563v1 Announce Type: cross Abstract: Software testing is critical for verifying that systems meet specified requirements, yet remains among the most time-consuming and expensive activitie

An Abstract Architecture for Explainable Autonomy in Hazardous Environments

SafetyDGX agent

arXiv:2606.07211v1 Announce Type: cross Abstract: Autonomous robotic systems are being proposed for use in hazardous environments, often to reduce the risks to human workers. In the immediate future,

Apple overhauls the App Store with subscription options for groups, businesses, and schools, retention messaging to reduce churn, bundled IAP reviews, and more (Andrew Orr/AppleInsider)

IndustryDGX agent

Andrew Orr / AppleInsider: Apple overhauls the App Store with subscription options for groups, businesses, and schools, retention messaging to reduce churn, bundled IAP reviews, and more — Apple is ov

Are Large Language Models Suitable for Graph Computation? Progress and Prospects

ResearchDGX agent

arXiv:2606.06865v1 Announce Type: new Abstract: Large language models (LLMs) have been increasingly explored for graph computation, where tasks require reasoning over structured relationships and algo

← Previous
1…125126127128129…168
Next →