AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,980 results
27 Jul 2026

Medical-Checklist: Assessing the Comprehension of Medical Images by Multimodal Models

Model ReleasesDGX agent

arXiv:2607.21998v1 Announce Type: new Abstract: This paper introduces a new benchmark test, Medical-Checklist, for assessing medical multimodal models. The recent advancements in multimodal models hav

Microsoft’s introduces its first agent-powered cybersecurity model

AgentsDGX agent

Microsoft Corp. today introduced its first in-house cybersecurity model, MAI-Cyber-1-Flash, and a companion agentic system called Project Perception that fields teams of artificial intelligence agents

Reliability Scales Inversely: Bigger Language Models Compound Mistakes Faster

SafetyDGX agent

arXiv:2607.18292v2 Announce Type: replace-cross Abstract: As language models scale, answers start truer but degrade faster: scaling buys capability but erodes reliability. The knowledge-gap account --


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Scaling Laws for Classical Machine Learning on Tabular Data: A Benchmark Study

Model ReleasesDGX agent

arXiv:2607.21866v1 Announce Type: new Abstract: Prior classical-ML learning-curve work fits power laws to tree, linear, and kernel models on tabular data, but at small scale: typically one curve, one

24 Jul 2026

OPTScientist: Multi-Agent Discovery of Typed Optimizer Programs for Transformer Pretraining

AgentsDGX agent

arXiv:2607.20486v1 Announce Type: new Abstract: Designing optimizers for modern deep learning remains a challenging scientific problem, requiring the joint consideration of optimization geometry, stat

The Human-AI Substitution Principle: When will you be replaced by AI in your organization?

ResearchDGX agent

arXiv:2607.20781v1 Announce Type: new Abstract: Artificial Intelligence (AI) is rapidly transforming organizations, raising a fundamental organizational and economic question: when will a human employ

23 Jul 2026

Announcing OpenWorker! An open-source agent that doesn't just chat with you, but delivers finished work -- like hand you a polished document…

Model ReleasesDGX agent

Announcing OpenWorker! An open-source agent that doesn't just chat with you, but delivers finished work -- like hand you a polished document, send a slack message, or update a calendar entry. Ask it t

Building multi-Region visualizations with Highcharts in Amazon Quick

TutorialsDGX agent

This post shows you how to build multi-Region carrier performance dashboards in Quick Sight using Highcharts custom visualizations to overcome native chart limitations. You will learn how to maintain

CreatiPoster: Towards Editable and Controllable Multi-Layer Graphic Design Generation

Model ReleasesDGX agent

arXiv:2506.10890v2 Announce Type: replace Abstract: Graphic design plays a crucial role in both commercial and personal contexts, yet creating high-quality, editable, and aesthetically pleasing graphi

Hypothesis-and-Refinement Learning of Organic Structures from Multimodal Spectroscopic Data

Model ReleasesDGX agent

arXiv:2607.19816v1 Announce Type: cross Abstract: Determining molecular structures from spectroscopic data remains fundamentally challenging because the inverse problem is intrinsically underdetermine

Minimize idle accelerators: Native RL job interleaving with co-operative time-slicing in llm-d

SafetyDGX agent

The math behind reinforcement learning (RL) post-training for large language models (LLMs) is notoriously unforgiving. As frontier AI labs push the boundaries of reasoning and coding models using RL p

OpenAI expands ChatGPT Health, which helps users with health-related queries and connects to services like Apple Health, to all logged-in US users over 18 (Ivan Mehta/TechCrunch)

IndustryDGX agent

Ivan Mehta / TechCrunch: OpenAI expands ChatGPT Health, which helps users with health-related queries and connects to services like Apple Health, to all logged-in US users over 18 — OpenAI said today

PaddlePaddle/HPD-Parsing · Hugging Face

Local AiDGX agent

HPD-Parsing: Hierarchical Parallel Document Parsing We introduce HPD-Parsing, a lightweight (1B) and high-throughput document parsing model built on a Hierarchical Parallel Decoding paradigm. Unified

21 Jul 2026

I thought this was a very good post by @c_valenzuelab and really highlights how transformative video models already are for the businesses t…

ApplicationsDGX agent

I thought this was a very good post by @c_valenzuelab and really highlights how transformative video models already are for the businesses that are using them. I think video models are one of the most

text/image-to-sim

AgentsDGX agent

Gizmo is a simulation-authoring agent that converts textual descriptions and reference images into structured, editable 3D scenes tailored for robotics workflows. It was publicly released as a beta on

20 Jul 2026

Accelerating automotive innovation with C4A-metal and Panasonic Automotive vSkipGen

HardwareDGX agent

As the automotive landscape accelerates toward software-defined vehicles, Cockpit Domain Controllers (CDCs) are becoming the core of next-generation in-cabin experiences. The ability to rapidly develo

16 Jul 2026

Automatic Ordinary Differential Equations Discovery For Biological Systems Using Large Language Model Powered Agentic System

AgentsDGX agent

arXiv:2607.13608v1 Announce Type: new Abstract: Automatic scientific discovery has long been a goal of computational scholars - a machine that can discover nature's secrets on its own, moving computat

Faithful Autoformalization of Natural Language Assertions

ResearchDGX agent

arXiv:2607.13303v1 Announce Type: cross Abstract: Formal contracts are essential for software testing and verification, yet writing them remains labor-intensive and error-prone. LLMs offer a promising

Reverse to Advance: Teleoperation-Cost Effective Hard Policy Learning from Reversed Easy Tasks

SafetyDGX agent

arXiv:2607.13455v1 Announce Type: new Abstract: High-quality teleoperation datasets are costly to collect, particularly for hard tasks. We observe that many tasks exhibit directional asymmetry: comple

Self-Improving AI Coding Agents Through Accumulated Behavioral Rules: A Closed-Loop Framework

AgentsDGX agent

arXiv:2607.13091v1 Announce Type: cross Abstract: LLM-based coding agents repeat the same classes of mistakes across sessions because they lack a mechanism to retain corrections from human review feed

15 Jul 2026

A Neurosymbolic Approach to Natural Language Formalization and Verification

SafetyDGX agent

arXiv:2511.09008v2 Announce Type: replace-cross Abstract: Large Language Models perform well at natural language interpretation and reasoning, but their lack of formal correctness guarantees limits th

Agentic systems for breast cancer treatment recommendations

Model ReleasesDGX agent

arXiv:2607.12051v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being explored for clinical decision support, but their reliability in complex oncology treatment planning

AI agents are already being used to improve the capabilities of our next-generation models. We believe with GPT-Red that we have started to …

SafetyDGX agent

AI agents are already being used to improve the capabilities of our next-generation models. We believe with GPT-Red that we have started to unlock a similar flywheel for safety, where today's models c

CrochetBench: Can Vision-Language Models Move from Describing to Doing in Crochet Domain?

Model ReleasesDGX agent

arXiv:2511.09483v3 Announce Type: replace Abstract: While multimodal large language models can describe visual content, their ability to generate executable procedures remains underexplored. CrochetBe

Good Benchmarks

ResearchDGX agent

arXiv:2607.12217v1 Announce Type: new Abstract: Good tasks are correct, solvable, verifiable, well-specified, and hard for interesting reasons. The best tasks describe a real problem an experienced pr

How Inference Compute Shapes Frontier LLM Evaluation

Model ReleasesDGX agent

arXiv:2606.17930v2 Announce Type: replace Abstract: AI evaluations are shifting toward harder tasks that benefit from longer trajectories involving tool use and iterative problem solving. As a result,

M2I2HA: Multi-modal Object Detection Based on Intra- and Inter-Modal Hypergraph Attention

SafetyDGX agent

arXiv:2601.14776v3 Announce Type: replace Abstract: Recent advances in multi-modal detection have significantly improved detection accuracy in challenging environments (e.g., low light, overexposure).

Same Loss, Same Noise, Opposite Schedules: Noise Structure and Optimizer Normalization Jointly Determine Whether Learning-Rate Cooldown Helps

ResearchDGX agent

arXiv:2607.12360v1 Announce Type: new Abstract: The cooldown phase of a warmup-stable-decay (WSD) learning-rate schedule, now a default in large-model pretraining, lowers the final training loss in so

SeqGPT: A Constrained Transformer Agent for the Inverse Designof Multi-Panel Composite Structures

Model ReleasesDGX agent

arXiv:2607.11910v1 Announce Type: cross Abstract: Optimizing composite stacking sequences to match continuous targets (e.g., Lamination or Buckling Parameters) with discrete manufacturing constraints

The Risk of Exposed Cloud Functions and How to Harden

Model ReleasesDGX agent

Written by: Corné de Jong Introduction Mandiant security assessments frequently identify publicly exposed serverless applications that lack authentication, often as a result of specific business requi

WanToFight: Real-Time Generative Game Engine for Multi-Player Combat Interaction

Local AiDGX agent

arXiv:2607.12592v1 Announce Type: new Abstract: We present WanToFight, a generative game engine that simulates real-time, two-player The King of Fighters '97 (KOF~'97) gameplay from keyboard input. Pr

14 Jul 2026

Entrust launches Agentic AI Trust Accelerator to move AI agents into production

Model ReleasesDGX agent

Identity-centered security solution company Entrust Corp. today launched the Agentic AI Trust Accelerator, a co-development program that brings together enterprises and technology partners to build th

Super impressed with what Harvinder and Suman have built at @airtap_ai. They've essentially turned SMS into a headless agentic execution lay…

AgentsDGX agent

Super impressed with what Harvinder and Suman have built at @airtap_ai. They've essentially turned SMS into a headless agentic execution layer for your mobile apps. You just text it to run errands, an

13 Jul 2026

When your brain works differently, AI isn’t a luxury—it’s accessibility

IndustryDGX agent

In this post, I share how AI serves as an accessibility tool for neurodivergent professionals. The system is built on Amazon Quick on your desktop, an AI-powered desktop and web assistant that compens

10 Jul 2026

FPGN: Redefining Ultra-Fast Programmable Gate-based Neural Acceleration with Differentiable LUTs

Local AiDGX agent

arXiv:2607.08427v1 Announce Type: cross Abstract: Achieving nanosecond-scale inference latency for deep neural networks (DNNs) has become a primary architectural concern for latency-critical applicati

GRE-Diff: Gaussian Room Embeddings for Structured Layout Diffusion

ResearchDGX agent

arXiv:2607.08086v1 Announce Type: new Abstract: Designing functional and aesthetically coherent floor plans requires exploring a vast space of possible room arrangements, a task that quickly becomes o

Self-Adaptive Anomaly Detection with Reinforcement Learning and Human Feedback in Connected Vehicles

AgentsDGX agent

arXiv:2607.08373v1 Announce Type: cross Abstract: Connected vehicles are autonomous cyber-physical systems whose behavior must be continuously monitored during operation to detect deviations from norm

9 Jul 2026

Agentic Data Environments

SafetyDGX agent

arXiv:2607.07397v1 Announce Type: new Abstract: Autonomous agents promise substantial gains in speed, scale, and labor efficiency, but their failures can impose abrupt and often irreversible costs. Th

AI Chatbot Suicide Risk Detection and Response: Human Validation Study of the Open-Source VERA-MH Safety Evaluation

Model ReleasesDGX agent

arXiv:2602.05088v4 Announce Type: replace Abstract: Millions of people now use generative AI chatbots for psychological support. Despite their promise, the most pressing question in AI for mental heal

Large Language Models (LLMs) and Generative AI in Cybersecurity and Privacy: A Survey of Dual-Use Risks, AI-Generated Malware, Explainability, and Defensive Strategies

Model ReleasesDGX agent

arXiv:2607.06963v1 Announce Type: cross Abstract: Large Language Models (LLMs) and generative AI (GenAI) systems, such as ChatGPT, Claude, Gemini, LLaMA, Copilot, Stable Diffusion by OpenAI, Anthropic

SpaCellAgent: A Self-Evolving LLM-Based Multi-Agent Framework for Trajectory Analysis

AgentsDGX agent

arXiv:2607.07467v1 Announce Type: new Abstract: Spatial and Single-cell transcriptomics are transformative in deciphering cellular dynamics. As the fundamental paradigm for reconstructing cell develop

8 Jul 2026

Improving LLM-Generated Process Model Quality Through Reinforcement Learning: The Role of Reward Function Design

Model ReleasesDGX agent

arXiv:2607.06175v1 Announce Type: cross Abstract: Large language models (LLMs) can generate BPMN process models from natural-language descriptions, yet supervised fine-tuning (SFT) limits their output

KernelEvolve: Scaling Agentic Kernel Coding for Heterogeneous AI Accelerators at Meta

SafetyDGX agent

arXiv:2512.23236v4 Announce Type: replace-cross Abstract: Making deep learning recommendation model (DLRM) training and inference fast and efficient is important. However, this presents three key syst

We raised 130M @ 1B for our series A To build the open superintelligence stack for everyone Pre-training concentrated frontier AI in a han…

HardwareDGX agent

We raised 130M @ 1B for our series A To build the open superintelligence stack for everyone Pre-training concentrated frontier AI in a handful of labs. RL changes who can build frontier AI and just wo

7 Jul 2026

3D Cal: An Open-Source Software Library for Depth Reconstruction on Vision-Based Tactile Sensors

ResearchDGX agent

arXiv:2511.03078v3 Announce Type: replace Abstract: Tactile sensing plays a key role in enabling dexterous and reliable robotic manipulation, but realizing this capability requires substantial calibra

CABTO: Context-Aware Behavior Tree Grounding for Robot Manipulation

ResearchDGX agent

arXiv:2603.16809v2 Announce Type: replace-cross Abstract: Behavior Trees (BTs) offer a powerful paradigm for designing modular and reactive robot controllers. BT planning, an emerging field, provides

Deep Learning-Based Characterization of Detonation-Cell Size Distributions in Soot-Foil Records

Model ReleasesDGX agent

arXiv:2607.03764v1 Announce Type: cross Abstract: The geometric size and regularity of detonation cells are key physical parameters for characterizing detonation waves. Traditional manual measurement

Flow-A11y: Flow-Aware Accessibility Testing

AgentsDGX agent

arXiv:2607.03100v1 Announce Type: cross Abstract: Modern web applications increasingly expose accessibility barriers through interaction flows rather than static page snapshots. Keyboard traps, focus

Paired Uterine Whole-Slide Images and Pathology Reports for Multimodal Computational Pathology

ResearchDGX agent

arXiv:2607.04020v1 Announce Type: new Abstract: Uterine diseases represent an important category of gynecologic pathology and require accurate histopathological assessment for diagnosis and treatment

Toward Trustworthy Large Language Model Agents in Healthcare

Model ReleasesDGX agent

arXiv:2607.05055v1 Announce Type: new Abstract: Healthcare appointment scheduling remains a persistent operational bottleneck, driven by manual coordination, fragmented legacy systems, and high admini

3 Jul 2026

AgenticDataBench: A Comprehensive Benchmark for Data Agents

Model ReleasesDGX agent

arXiv:2607.01647v1 Announce Type: cross Abstract: Data science aims to derive actionable insights from heterogeneous raw data, unlocking the value of the massive amounts of data generated in modern so

Don't train the model, evolve the harness. I read a brilliant blog post from Hugging Face where they took a frozen open model scoring 0% on …

Model ReleasesDGX agent

Don't train the model, evolve the harness. I read a brilliant blog post from Hugging Face where they took a frozen open model scoring 0% on a hard legal agent benchmark, left its weights alone, and le

Enerzyme: A Framework for Efficient Training of Reactive Neural Network Potentials for Enzyme Catalysis with Application to Methyltransferases

TutorialsDGX agent

arXiv:2607.01362v1 Announce Type: cross Abstract: Quantum mechanical (QM) cluster models provide an effective framework for mechanistic studies of enzymatic reactions but remain computationally demand

Exploring Large Language Models for Access Control Policy Synthesis and Summarization

SafetyDGX agent

arXiv:2510.20692v2 Announce Type: replace-cross Abstract: Cloud computing is ubiquitous, with a growing number of services being hosted on the cloud every day. Typical cloud compute systems allow admi

Something about this year’s @aiDotEngineer World’s Fair just hit different. Last year was the year of “let the agents rip.” This year was th…

Model ReleasesDGX agent

Something about this year’s @aiDotEngineer World’s Fair just hit different. Last year was the year of “let the agents rip.” This year was the year of realizing that autonomy without structure creates

TestEvo-Bench: An Executable and Live Benchmark for Test and Code Co-Evolution

Model ReleasesDGX agent

arXiv:2607.02469v1 Announce Type: cross Abstract: Software tests and code evolve together: a code change should be followed by new or updated tests that record the new software behavior. Yet existing

2 Jul 2026

Benchmarking Frontier LLMs on Arabic Cultural and Sociolinguistic Knowledge: A Cross-Evaluation Framework with Human SME Ground Truth

Model ReleasesDGX agent

arXiv:2607.00139v1 Announce Type: new Abstract: The cost of human expert evaluation is a principal bottleneck to deploying language models in specialized, high-stakes domains. This is particularly acu

1 Jul 2026

Beyond Static Prompts: Building Scale-Proof, Polymorphic Multi-Agent Systems with Google's ADK

Model ReleasesDGX agent

As enterprise generative AI transitions from simple, conversational chatbots to autonomous multi-agent workflows, developers face a critical bottleneck: scale. In a production environment, an enterpri

GUIDE: Resolving Domain Bias in GUI Agents through Real-Time Web Video Retrieval and Plug-and-Play Annotation

SafetyDGX agent

arXiv:2603.26266v3 Announce Type: replace Abstract: Large vision-language models have endowed GUI agents with strong general capabilities for interface understanding and interaction. However, due to i

Locker-based Truck-Drone Routing with Integrated Considerations of Pickups, Deliveries, and No-Fly Zones

SafetyDGX agent

arXiv:2606.30680v1 Announce Type: cross Abstract: Truck-drone delivery is an emerging last-mile logistics mode combining the long-haul capacity of trucks with the flexible service capability of drones

← Previous
1…3637383940…83
Next →