AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,980 results
29 May 2026

From petabytes to predictions: Easy BigQuery insights in Google Sheets

TutorialsDGX agent

Many organizations’ single source of truth is data that resides in BigQuery, Google’s governed, secure and petabyte-scale data platform. However, the 'last mile' of ad-hoc analysis, modeling, and repo

GenesisFunc: Multi-Agent Data Generation for Accurate and Generalizable Function-Calling

AgentsDGX agent

arXiv:2605.28835v1 Announce Type: cross Abstract: Large Language Models (LLMs) extend their capabilities through function-calling (FC), which relies on training data with high quality, diversity, and

GTA: Generating Long-Horizon Tasks for Web Agents at Scale

Model ReleasesDGX agent

arXiv:2605.29218v1 Announce Type: new Abstract: Web agents, which couple language models with browsing and tool-use capabilities, show promise as open web assistants. Yet progress is increasingly limi


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Hallucination Detection-Guided Preference Optimization for Clinical Summarization

Model ReleasesDGX agent

arXiv:2605.28910v1 Announce Type: cross Abstract: Large language models (LLMs) have shown promise on summarization tasks, but they often produce hallucinations, which are unsupported or incorrect stat

Improving agents The old way: Manually reading traces, looking for patterns, writing evals, and creating fixes. The better way: Letting Lang…

AgentsDGX agent

This post from LangChain's Harrison Chase contrasts traditional manual methods of improving AI agents (tracing execution, identifying patterns, writing evaluations, and implementing fixes) with a more

InsightEval: An Expert-Curated Benchmark for Assessing Insight Discovery in LLM-Driven Data Agents

Model ReleasesDGX agent

arXiv:2511.22884v2 Announce Type: replace Abstract: Data analysis has become an indispensable part of scientific research. To discover the latent knowledge and insights hidden within massive datasets,

Jailbreaking and Mitigation of Vulnerabilities in Large Language Models

SafetyDGX agent

arXiv:2410.15236v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have transformed artificial intelligence by advancing natural language understanding and generation, enabling app

KLAS: Using Similarity to Stitch Neural Networks for Improved Accuracy-Efficiency Tradeoffs

ResearchDGX agent

arXiv:2605.29259v1 Announce Type: cross Abstract: Given the wide range of deployment targets, flexible model selection is essential for optimizing performance within a given compute budget. Recent wor

Low-Magnification SEM May Suffice: Interpretable Deep Learning for Multi-Scale Fracture-Cause Classification in Zirconia-Toughened Alumina

SafetyDGX agent

arXiv:2605.29798v1 Announce Type: new Abstract: Reliable identification of fracture origins in alumina matrix composite hip and knee implants is critical for quality assurance and patient safety, yet

My MLSys keynote on AI writing systems code got more interest than I expected. The recording will take a while, so in the finest tradition o…

TutorialsDGX agent

My MLSys keynote on AI writing systems code got more interest than I expected. The recording will take a while, so in the finest tradition of AI labs sharing blog posts, we’re starting the Core Automa

Opt-Verifier: Unleashing the Power of LLMs for Optimization Modeling via Dual-Side Verification

ResearchDGX agent

arXiv:2605.29556v1 Announce Type: new Abstract: Building mathematical optimization models is critical in operations research (OR), while it requires substantial human expertise. Recent advancements ha

OptSkills: Learning Generalizable Optimization Skills from Problem Archetypes via Cluster-Based Distillation

Model ReleasesDGX agent

arXiv:2605.29829v1 Announce Type: new Abstract: Leveraging Large Language Models (LLMs) to automatically formulate and solve optimization problems from natural language has emerged as an efficient par

PassNet: Scaling Large Language Models for Graph Compiler Pass Generation

ApplicationsDGX agent

arXiv:2605.29357v1 Announce Type: new Abstract: Modern tensor compilers such as TorchInductor deliver substantial speedups on mainstream models, yet face a systematic performance ceiling on long-tail

Projectional Decoding: Towards Semantic-Aware LLM Generation

ResearchDGX agent

arXiv:2605.30054v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to generate software artifacts across many software engineering (SE) tasks, yet ensuring the semant

ProjectionBench: Evaluating Scientific Hypothesis Generation in LLMs Under Progressive Information Disclosure

Model ReleasesDGX agent

arXiv:2605.30284v1 Announce Type: new Abstract: Scientific discovery is an inherently creative and uncertain process, requiring reasoning beyond the recall of known knowledge. While many benchmarks ha

Rethinking Post-Training Recipes for Multimodal Time-Series Forecasting

Model ReleasesDGX agent

arXiv:2605.29401v1 Announce Type: new Abstract: Time-Series Foundation Models (TSFMs) excel at zero-shot unimodal forecasting using numerical data, but unlike LLMs they cannot consume multimodal, non-

RoboWits: Unexpected Challenges for Robotic Creative Problem Solving

Model ReleasesDGX agent

arXiv:2605.30326v1 Announce Type: cross Abstract: The ability to reason, adapt, and creatively solve problems under unexpected challenges is essential for robots operating in real-world environments.

SCOPE: A Lightweight-training LLM Framework for Air Traffic Control Readback Monitoring

ResearchDGX agent

arXiv:2605.29543v1 Announce Type: cross Abstract: Pilot readback of Air Traffic Control (ATC) voice instructions is a primary safeguard against miscommunication in air transportation. However, readbac

Shame on you Apple, sending @mcuban’s thoughtful email to spam. Seriously? Glad I have learned not to trust your algorithm.

SafetyDGX agent

Gary Marcus criticizes Apple's email filtering system for incorrectly routing Mark Cuban's message to spam, highlighting a failure in Apple's spam detection algorithm. The post reflects concerns about

TANDEM: Temporal-Aware Neural Detection for Multimodal Hate Speech

Model ReleasesDGX agent

arXiv:2601.11178v2 Announce Type: replace Abstract: Social media platforms are increasingly dominated by long-form multimodal content, where harmful narratives are constructed through a complex interp

The Anatomy of Conversational Scams: A Topic-Based Red Teaming Analysis of Multi-Turn Interactions in LLMs

SafetyDGX agent

arXiv:2601.03134v2 Announce Type: replace Abstract: As LLMs gain persuasive capabilities through extended dialogues, they create new opportunities for studying adversarial conversational behavior in e

ViASNet: A Video Ad Saliency Network for Predicting Dynamic Saliency and Viewer Engagement

ResearchDGX agent

arXiv:2605.29302v1 Announce Type: new Abstract: The digital media landscape has seen a pervasive shift toward short-form video advertising on TV, social media and e-commerce platforms. The present stu

28 May 2026

A Matter of TASTE: Improving Coverage and Difficulty of Agent Benchmarks

Model ReleasesDGX agent

arXiv:2605.28556v1 Announce Type: new Abstract: As agent capabilities advance, existing benchmarks, such as au^2-Bench, are becoming increasingly saturated. Yet constructing new benchmark tasks remain

A Patient-Specific Pulmonary Arterial Tree Digital Twin to Extract Pulmonary Embolism Biomarkers

ResearchDGX agent

arXiv:2605.28217v1 Announce Type: new Abstract: Pulmonary embolism, the obstruction of a pulmonary artery by a blood clot, is one of the leading causes of acute cardiovascular syndrome. In clinical pr

Ariel-ML: Computing Parallelization with Embedded Rust for Neural Networks on Heterogeneous Multi-core Microcontrollers

Model ReleasesDGX agent

arXiv:2512.09800v2 Announce Type: replace Abstract: Low-power microcontroller (MCU) hardware is currently evolving from single-core architectures to predominantly multi-core architectures. In parallel

Asana acquires StackAI, a no-code platform for building AI agents, for 75M as part of Asana's broader AI pivot; PitchBook: StackAI raised ~20M (Russell Brandom/TechCrunch)

AgentsDGX agent

Russell Brandom / TechCrunch: Asana acquires StackAI, a no-code platform for building AI agents, for 75M as part of Asana's broader AI pivot; PitchBook: StackAI raised ~20M — Asana has acquired the wo

Asana acquires StackAI to run AI agent workflows across enterprise systems

AgentsDGX agent

Work management software company Asana Inc. today said it has completed the acquisition of StackAI Inc., a no-code platform for building artificial intelligence agents, in a deal that adds the ability

AssertLLM2: A Comprehensive LLM Benchmark for Assertion Generation from Design Specifications

Model ReleasesDGX agent

arXiv:2605.27472v1 Announce Type: cross Abstract: Assertion-based verification (ABV) is a cornerstone of modern hardware design, yet manually translating design intent into formal SystemVerilog Assert

Backdoor Attacks on Fault Detection and Localization in Cyber-Physical Systems

Local AiDGX agent

arXiv:2605.27674v1 Announce Type: cross Abstract: Cyber-Physical Systems (CPS) integrate sensing, communication, computation, and control to support critical infrastructure, including smart grids, ind

Breaking the Script Barrier: Enabling Automatic Alignment for PoS-based ASR Error Analysis in Non-Latin Scripts

SafetyDGX agent

arXiv:2605.28438v1 Announce Type: new Abstract: Automatic Speech Recognition (ASR) systems are commonly evaluated using aggregate metrics such as Word Error Rate (WER), which do not capture the lingui

Bridging the Stability-Expressivity Gap: Synthetic Data Scaling and Preference Alignment for Low-Resource Spoken Language Models

Model ReleasesDGX agent

arXiv:2605.27383v1 Announce Type: cross Abstract: Spoken Language Models (SLMs) have emerged as a promising paradigm for speech synthesis by bypassing explicit grapheme-to-phoneme pipelines. However,

CLEAR-NeRF: Collinearity and Local-region Enhanced Accurate 3D Reconstruction in Unbounded Scenes

TutorialsDGX agent

arXiv:2605.28125v1 Announce Type: new Abstract: Many real-world 3D reconstruction applications demand photorealism and metric accuracy across unbounded, complex scenes with challenging lighting and im

Cyberbullying Governance on Social Media: A Unified Framework from Content Identification to Intervention

SafetyDGX agent

arXiv:2605.27584v1 Announce Type: new Abstract: The proliferation of social media platforms and online communities has inadvertently catalyzed the spread of cyberbullying, hate speech, and other forms

DeltaMCP: Incremental Regeneration via Spec-Aware Transformation for MCP servers

AgentsDGX agent

arXiv:2605.28148v1 Announce Type: cross Abstract: The rapid development of LLMs coupled with the introduction of Model Context Protocol (MCP) has revolutionized how intelligent agents interact with AP

Eliot: Interactively nderline{E}xploring Fast-Changing Scientific nderline{Li}terature Trends with nderline{O}nline Danderline{t}a and Learning

ResearchDGX agent

arXiv:2605.27610v1 Announce Type: cross Abstract: The rapid growth of scientific publishing has made it increasingly difficult to track how fast-moving areas evolve. Search engines and LLM-based assis

Enhancing Trustworthy GUI Grounding via Self-Critiqued Reinforcement Learning

Model ReleasesDGX agent

arXiv:2510.27266v2 Announce Type: replace Abstract: Autonomous graphical user interface (GUI) agents rely on accurate GUI grounding, which maps language instructions to on-screen coordinates, to execu

Fine-Tuning Vision-Language Models for Understanding Current Damage and Scoring Priority with Quality Guard Agent

AgentsDGX agent

arXiv:2605.27452v1 Announce Type: new Abstract: Bridge inspection in Japan requires mandatory visual assessments every five years, yet qualitative damage ratings (levels a-e) assigned by different eng

How Endava builds an agentic organization with Codex

AgentsDGX agent

Endava, a software services company, leverages OpenAI's Codex to transform its organizational operations into an agentic model where AI agents autonomously handle tasks and decision-making. The approa

How the University of Central Oklahoma is using AI to streamline analysis of complex criminal cases

Model ReleasesDGX agent

In the high-stakes world of forensic science, time is the enemy of justice. The University of Central Oklahoma (UCO) Forensic Science Institute (FSI) was looking for an innovative AI solution that cou

Human-AI Collaboration for Estimating Scientific Replicability

ResearchDGX agent

arXiv:2605.27394v1 Announce Type: cross Abstract: Determining whether published scientific findings can successfully be replicated is a long-standing challenge in the empirical sciences. Existing appr

JECA^2: Judgment-Explanation Consistent Adversarial Attack against Forensic Vision-Language Models

SafetyDGX agent

arXiv:2605.28609v1 Announce Type: new Abstract: Forensic vision-language models (VLMs) have recently been developed to detect image tampering and provide natural-language explanations. However, their

join us for a behind the curtains look at LangSmith Engine (our agent that helps make your agent better)

AgentsDGX agent

join us for a behind the curtains look at LangSmith Engine (our agent that helps make your agent better) One of the best ways to learn what LangSmith Engine is capable of is to talk to the team that b

Learning High-Dimensional Parity Functions with Product Networks using Gradient Descent

SafetyDGX agent

arXiv:2605.28612v1 Announce Type: new Abstract: Parity functions are fundamental Boolean operations with critical applications across machine learning, cryptography, and error correction. Yet, learnin

Mahalanobis PatchCore: Covariance-Aware and Streaming-Compatible Industrial Anomaly Detection

Model ReleasesDGX agent

arXiv:2605.27748v1 Announce Type: cross Abstract: Industrial visual anomaly detection is usually one-class: normal images are abundant, while defects are rare, heterogeneous, and often unavailable dur

Mobile-Aptus: Confidence-Driven Proactive and Robust Interaction in MLLM-based Mobile-Using Agents

SafetyDGX agent

arXiv:2605.28629v1 Announce Type: new Abstract: Recent advancements in multimodal large language models (MLLMs) have shown exceptional potential in enabling mobile-using agents to autonomously execute

NL-MambaXCT: Self-Supervised Nested-Learning Mamba for Nomex Honeycomb X-ray CT Defect Classification

Model ReleasesDGX agent

arXiv:2605.27454v1 Announce Type: cross Abstract: X-ray computed tomography (XCT) is widely used for non-destructive testing of Nomex honeycomb structures in aerospace manufacturing, but industrial in

OralAgent: Integrating Reasoning, Tools, and Knowledge for Interactive Dental Image Analysis

Model ReleasesDGX agent

arXiv:2605.27378v1 Announce Type: new Abstract: Dental image analysis plays a pivotal role in supporting accurate diagnosis and treatment planning in oral healthcare. Although recent advances have pro

Persuade Me if You Can: A Framework for Evaluating Persuasion Effectiveness and Susceptibility Among Large Language Models

Model ReleasesDGX agent

arXiv:2503.01829v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) demonstrate persuasive capabilities that rival human-level persuasion. While these capabilities can be used for s

PrimitiveVLA: Learning Reusable Motion Primitives for Efficient and Generalizable Robotic Manipulation

TutorialsDGX agent

arXiv:2605.28634v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models offer a promising paradigm for generalist robotic policies, yet their adaptation is hindered by data inefficiency an

TCP-MCP: Landscape-Guided Co-Evolution of Prompts and Communication Topologies for Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2605.27850v1 Announce Type: new Abstract: Effective multi-agent systems cannot be designed by selecting prompts or communication graphs in isolation. Agent behavior depends on the information an

Technical Report: Exploring the Emerging Threats of the Agent Skill Ecosystem

AgentsDGX agent

arXiv:2605.28588v1 Announce Type: cross Abstract: We analyzed 3,984 AI agent skills from major marketplaces and found 76 confirmed malicious payloads, including credential theft, backdoor installation

The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic Video Generation

Model ReleasesDGX agent

arXiv:2601.17737v3 Announce Type: replace-cross Abstract: Recent advances in video generation have produced models capable of synthesizing stunning visual content from simple text prompts. However, th

Tree of Thoughts as a Classical Heuristic Search Problem: Formal Foundations and Design Patterns

ResearchDGX agent

arXiv:2605.28566v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable reasoning capabilities, yet their standard generation process -- auto-regressive token predict

UniMaia: Steering Chess Policies with Language for Human-like Play

Model ReleasesDGX agent

arXiv:2605.27767v1 Announce Type: cross Abstract: Recent advances in large language models have enabled natural language to serve as a flexible interface for controlling complex systems, but often at

27 May 2026

6/ More details on our business and the rise of autonomous software engineering: https://cognition.ai/blog/series-d

AgentsDGX agent

Cognition AI announced details about their Series D funding round, focusing on developments in autonomous software engineering capabilities. The post likely discusses their business progress, product

A glimpse of an exciting future where tax professionals will be able to spend more time advising customers and explaining the tax returns!

ToolsDGX agent

A glimpse of an exciting future where tax professionals will be able to spend more time advising customers and explaining the tax returns! At @ThriveHoldings, we built a product with @OpenAI to automa

A Universal Cliff and a Design Fingerprint: Cross-Section Defect Detection Under LLM Orchestration

SafetyDGX agent

arXiv:2605.26174v1 Announce Type: cross Abstract: Production language-model systems answer a request by partitioning it across an invisible orchestration of worker agents that recompose one integrated

AssetGen: Deployable 3D Asset Generation at Interactive Speed

HardwareDGX agent

arXiv:2605.26137v1 Announce Type: cross Abstract: While 3D generation is progressing rapidly, recent work has often focused on obtaining high-resolution assets, leaving user experience and deployabili

Athena: Enhancing Multimodal Reasoning with Data-efficient Process Reward Models

SafetyDGX agent

arXiv:2506.09532v5 Announce Type: replace-cross Abstract: We present Athena-PRM, a multimodal process reward model (PRM) designed to evaluate the reward score for each step in solving complex reasonin

AutoDFT: A Closed-Loop Multi-Agent Framework for Autonomous DFT Calculations

Model ReleasesDGX agent

arXiv:2605.26179v1 Announce Type: cross Abstract: Density functional theory (DFT) serves as the basis for computational discovery in materials science and chemistry, yet each calculation demands exten

← Previous
1…6263646566…83
Next →