AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,979 results
Model Releases

🆕This Year In Claude https://www.youtube.com/watch?v=uU5Gv2h8-9g @simonw chats with @_catwu and @trq212 about the state of: - @claudeai Cod…

DGX agent

🆕This Year In Claude https://www.youtube.com/watch?v=uU5Gv2h8-9g @simonw chats with @_catwu and @trq212 about the state of: - @claudeai Code - Claude Fable - @anthropicai culture & product strategy -

model-releasesswyx--x
15 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Visual Species Recognition with Large Multimodal Models as Post-Hoc Correctors

DGX agent

arXiv:2512.15748v2 Announce Type: replace-cross Abstract: Visual Species Recognition (VSR) is a fundamental task in scientific disciplines that require species-level identification, including ecology,

researcharxiv-cs-cv
15 Jul 2026
Local Ai

Local AI is the future. Learning how to run Opensource models (Inference), how to evaluate them systematically (Evals), and how to customize…

DGX agent

Local AI is the future. Learning how to run Opensource models (Inference), how to evaluate them systematically (Evals), and how to customize them (Fine-tuning / RL / Post-training) are invaluable skil

local-aiclem-delangue--x
14 Jul 2026
Hardware

Post-Train NVIDIA Cosmos 3 in One Day Using Agent Skills

DGX agent

NVIDIA Cosmos 3 was post‑trained in under a day using TAO agent skills and LoRA adapters, raising accuracy on the Woven Traffic Safety video QA dataset from 54.41 % to 93.35 %. The mixture‑of‑transfor

hardwarenvidia-developer
14 Jul 2026
Tutorials

Scaling UX testing with Amazon Nova Act: A new approach to user flow analysis

DGX agent

Using generative AI enables parallel execution of comprehensive user flow testing at scale. This solution demonstrates how to build a cloud-deployed UX testing platform that automatically generates te

tutorialsaws-ml-blog
14 Jul 2026
Agents

WANDR tests how well agents discover large sets of entities and verify specific facts about each one. It provides a dense, interpretable eva…

DGX agent

WANDR tests how well agents discover large sets of entities and verify specific facts about each one. It provides a dense, interpretable eval signal that reveals whether an agent fails, and where. The

agentsperplexity--x
14 Jul 2026
Model Releases

Key findings from the 2026 Public Sector M-Trends report and beyond

DGX agent

In 2026, the public sector is no longer defending a traditional perimeter. Instead, they are defending a complex web of interconnected trust relationships against adversaries that now operate at machi

model-releasesgoogle-cloud-ai
13 Jul 2026
Safety

Absolutely fascinating work by @SakanaAILabs reproducing @kenneth0stanley Picbreeder in a non-interactive, VLM-agentic way. I've had years t…

DGX agent

Absolutely fascinating work by @SakanaAILabs reproducing @kenneth0stanley Picbreeder in a non-interactive, VLM-agentic way. I've had years to reflect on Kenneth Stanley's ideas as originally communica

safetydavid-ha--x
11 Jul 2026
Agents

Agentic AI and Retrieval-Augmented Models in Straight-Through Underwriting

DGX agent

arXiv:2607.07858v1 Announce Type: new Abstract: Artificial intelligence (AI) is beginning to reshape actuarial practice, particularly in domains that require reasoning over unstructured documents, het

agentsarxiv-cs-ai
10 Jul 2026
Applications

An Online Reference-Free Evaluation Framework for Flowchart Image-to-Code Generation

DGX agent

arXiv:2602.13376v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly used in document processing pipelines to convert flowchart images into structured code (e.g., M

applicationsarxiv-cs-ai
10 Jul 2026
Hardware

ARGUS: Accelerated, Robust, General, and Unsupervised Cell Tracking Solutions

DGX agent

arXiv:2607.08297v1 Announce Type: new Abstract: Background and Objective: Quantitative analysis of cell dynamics is central to modern biological research, providing critical insights into immune cell

hardwarearxiv-cs-cv
10 Jul 2026
Model Releases

Blind-Spots-Bench: Evaluating Blind Spots in Multimodal Models

DGX agent

arXiv:2607.08317v1 Announce Type: new Abstract: Modern AI models achieve strong performance on many established benchmarks, yet they still fail on tasks that humans find almost trivial, such as manipu

model-releasesarxiv-cs-ai
10 Jul 2026
Research

Evaluating the Effect of Frame Rate in Sequence-Based Classification of Autism-Related Self-Stimulatory Hand Idiosyncrasies

DGX agent

arXiv:2607.07957v1 Announce Type: new Abstract: Autism spectrum disorder (ASD) affects over 75 million individuals worldwide, yet scalable computational methods for remote behavioral screening remain

researcharxiv-cs-ai
10 Jul 2026
Agents

From Legacy Documentation to OSCAL: An MCP-Based Agent Pipeline for Threat-Informed Continuous Compliance in Critical Infrastructure

DGX agent

arXiv:2607.08288v1 Announce Type: cross Abstract: In critical infrastructure, operational technology environments often cannot be actively scanned, and yet active system feedback is needed for risk as

agentsarxiv-cs-ai
10 Jul 2026
Agents

Game Theory Driven Multi-Agent Framework Mitigates Language Model Hallucination

DGX agent

arXiv:2607.08403v1 Announce Type: new Abstract: The application of lightweight Large Language Models in rule-based scientific domains remains severely limited by their tendency to mimic linguistic pat

agentsarxiv-cs-ai
10 Jul 2026
Applications

GradInf: Gradient Estimation as Probabilistic Inference

DGX agent

arXiv:2607.07840v1 Announce Type: cross Abstract: Gradient estimation -- the task of computing the gradient of the expected value of a probabilistic program -- has diverse applications in scientific c

applicationsarxiv-cs-lg
10 Jul 2026
Model Releases

IMProofBench: Benchmarking AI on Research-Level Mathematical Proof Generation

DGX agent

arXiv:2509.26076v2 Announce Type: replace Abstract: As the mathematical capabilities of large language models (LLMs) improve, it becomes increasingly important to evaluate their performance on researc

model-releasesarxiv-cs-cl
10 Jul 2026
Safety

In vivo feasibility study of humanoid robots in surgery

DGX agent

arXiv:2607.07972v1 Announce Type: new Abstract: Recent advances in actuation, control and learning have rapidly pushed humanoid robots from a distant vision towards near-term real-world deployment. He

safetyarxiv-cs-ro
10 Jul 2026
Safety

PLURAL: A Global Dataset for Value Alignment

DGX agent

arXiv:2607.08034v1 Announce Type: cross Abstract: Large language models (LLMs) are used worldwide, yet disproportionately reflect Western values, limiting their ability to represent diverse value syst

safetyarxiv-cs-ai
10 Jul 2026
Safety

Search-based Testing of Vision Language Models for In-Car Scene Understanding

DGX agent

arXiv:2607.02300v2 Announce Type: replace Abstract: In the automotive domain, in-car scene understanding (ISU) enables the detection of safety-critical events, such as driver distraction, and supports

safetyarxiv-cs-cv
10 Jul 2026
Agents

Sierra isn't the first to build this - Ramp, Stripe, CoinBase also have If you want an open source version - check out OpenSWE: https://gith…

DGX agent

Sierra isn't the first to build this - Ramp, Stripe, CoinBase also have If you want an open source version - check out OpenSWE: https://github.com/langchain-ai/open-swe We use it internally (mostly fo

agentsharrison-chase--x
10 Jul 2026
Research

Systematic Evaluation of Learning Rate Scheduling Strategies Across Heterogeneous Architectures

DGX agent

arXiv:2607.08511v1 Announce Type: cross Abstract: Choosing a learning rate scheduling strategy is critical to neural network training, but manual selection is costly and rarely exhaustive. While class

researcharxiv-cs-cv
10 Jul 2026
Safety

TNODEV: Toolbox for Neural ODE Verification

DGX agent

arXiv:2606.16567v2 Announce Type: replace Abstract: Neural ordinary differential equations (neural ODE) gained attention in safety critical settings such as continuous-time controllers for cyber-physi

safetyarxiv-cs-ai
10 Jul 2026
Industry

Two underappreciated features in Codex: - the ability to steer mid-process - computer use I don't think labs realize how important computer …

DGX agent

This post discusses two underutilized capabilities of OpenAI's Codex: mid-process steering (the ability to guide code generation in real-time during execution) and computer use functionality. The auth

industryallie-k--miller--x
10 Jul 2026
Safety

XALPHA: A Memory-Driven AI Quant Researcher for Hypothesis-to-Code Alpha Discovery

DGX agent

arXiv:2607.08332v1 Announce Type: new Abstract: Financial markets are noisy, non-stationary, and high-dimensional, making it difficult to discover predictive and robust trading signals. Alpha discover

safetyarxiv-cs-cl
10 Jul 2026
Research

A Multi-Analyst LLM Pipeline for Auditable Rule Discovery Across 68 Public Physiological Corpora

DGX agent

arXiv:2607.06802v1 Announce Type: cross Abstract: Open physiological corpora are heterogeneous: they use different sensors, labels, sampling rates, recording settings, and clinical endpoints. They can

researcharxiv-cs-ai
9 Jul 2026
Research

CEVAR: Centerline Embedding Extraction for Endovascular Aneurysm Repair

DGX agent

arXiv:2606.15667v2 Announce Type: replace Abstract: Long-term mortality rates after endovascular aneurysm repair (EVAR) remain elevated due to post-EVAR rupture caused by loss of seal in stent graft s

researcharxiv-cs-cv
9 Jul 2026
Hardware

DYNA-PRUNER: Input-Adaptive Data-Model Co-Pruning for Efficient and Scalable Spatio-Temporal Media Prediction

DGX agent

arXiv:2606.15346v2 Announce Type: replace Abstract: Spatio-temporal prediction supports radar/satellite nowcasting and city-scale traffic monitoring, but modern models are often too expensive for real

hardwarearxiv-cs-cv
9 Jul 2026
Local Ai

Enhancing enterprise inference on Amazon SageMaker HyperPod with data capture, Hugging Face, NVMe, and Route 53 integration

DGX agent

In this post, we walk through five capabilities now available in SageMaker HyperPod inference: multi-tier data capture for auditing and model improvement, direct deployment from Hugging Face Hub, loca

local-aiaws-ml-blog
9 Jul 2026
Model Releases

Evaluating SageMath-Augmented LLM Agents for Computational and Experimental Mathematics

DGX agent

arXiv:2607.06820v1 Announce Type: new Abstract: Recent advances in AI for Mathematics have focused largely on autoformalization and theorem proving, leaving the role of Computer Algebra Systems (CAS)

model-releasesarxiv-cs-ai
9 Jul 2026
Agents

From Agentic to Autogenic Network Management for AI-Native 6G and Beyond: A Standards Perspective

DGX agent

arXiv:2607.06786v1 Announce Type: cross Abstract: Standards bodies, including TM Forum, 3GPP, and ETSI, are converging on Agentic AI as the foundation for next-generation network management, where Lar

agentsarxiv-cs-ai
9 Jul 2026
Model Releases

From Text to Parameters: Predicting Item Parameters from Embedding Regularization with Reliability and Design Ceilings

DGX agent

arXiv:2607.07141v1 Announce Type: new Abstract: Newly developed items must ordinarily be field tested before their psychometric properties are known, creating a cold start problem for item calibration

model-releasesarxiv-cs-cl
9 Jul 2026
Research

LEMUR 2: Unlocking Neural Network Diversity for AI

DGX agent

arXiv:2607.06839v1 Announce Type: cross Abstract: Existing NAS benchmarks (e.g., NAS-Bench, NATS-Bench) cover only narrow, task-specific regions of the architectural design space and lack cross-domain

researcharxiv-cs-cv
9 Jul 2026
Model Releases

MedPMC: A Systematic Framework for Scaling High-Fidelity Medical Multimodal Data for Foundation Models

DGX agent

arXiv:2607.07673v1 Announce Type: new Abstract: Medicine is inherently multimodal, requiring clinicians to synthesize information across diverse data streams. Yet the development of multimodal foundat

model-releasesarxiv-cs-cv
9 Jul 2026
Model Releases

Optimization-Embedded Active Multi-Fidelity Surrogate Learning for Multi-Condition Airfoil Shape Optimization

DGX agent

arXiv:2603.17057v2 Announce Type: replace-cross Abstract: Active multi-fidelity surrogate modeling is developed for multi-condition airfoil shape optimization to reduce high-fidelity CFD cost while re

model-releasesarxiv-cs-lg
9 Jul 2026
Research

Pixel-Precise Explainable Stress Indexing: A Semantic Segmentation Framework for Disease Severity Quantification in Field Crops

DGX agent

arXiv:2607.06585v1 Announce Type: new Abstract: Plant diseases, resulting from both biotic and abiotic stresses, cause an estimated 20-40% loss in global agricultural yield annually, resulting in econ

researcharxiv-cs-cv
9 Jul 2026
Model Releases

Rethinking Multimodal Time-Series Forecasting Evaluation

DGX agent

arXiv:2607.06973v1 Announce Type: new Abstract: We introduce a new context-enriched, multimodal time series forecasting benchmark, TimesX. TimesX contains a wide selection of high-quality real-world t

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

Safely run AI-generated code in Cloud Run sandboxes

DGX agent

Here’s a question we hear often at Google Cloud: How do you safely run AI-generated code or untrusted binaries without putting your host application, data, and cloud credentials at risk? In other word

model-releasesgoogle-cloud-ai
9 Jul 2026
Industry

Socialists imagine a class struggle. In their made-up fantasy the CEO is in competition with low level workers, the wealthy entrepreneur is …

DGX agent

Socialists imagine a class struggle. In their made-up fantasy the CEO is in competition with low level workers, the wealthy entrepreneur is stealing from the underpaid nurse. In reality, workers do no

industryelon-musk--x
9 Jul 2026
Model Releases

Solve harder problems with AlphaEvolve, now available to everyone on Google Cloud

DGX agent

Many of the most challenging and valuable problems in the world are related to optimization. Now, AI is now making these problems tractable. If you've ever tried to design a microchip, plan a delivery

model-releasesgoogle-cloud-ai
9 Jul 2026
Model Releases

SpaceXAI's Grok 4.5 takes the #1 spot on AutomationBench-AA with a score of 51%, ahead of Claude Fable 5 (49%) and Claude Opus 4.8 (48%) at …

DGX agent

SpaceXAI's Grok 4.5 takes the #1 spot on AutomationBench-AA with a score of 51%, ahead of Claude Fable 5 (49%) and Claude Opus 4.8 (48%) at roughly a quarter of their cost per task - the first model t

model-releaseselon-musk--x
9 Jul 2026
Safety

TACoS: Weakly Supervised Learning of Two-Dimensional Materials from Scribble Annotations to Precise Segmentation

DGX agent

arXiv:2607.07169v1 Announce Type: new Abstract: The precise pixel-level localization of 2D material flakes is crucial for high-throughput screening. However, traditional fully supervised methods rely

safetyarxiv-cs-cv
9 Jul 2026
Research

Video-Based Detection of squint and cataract for accessibility-aware adaptive web interface rendering

DGX agent

arXiv:2607.07099v1 Announce Type: new Abstract: Squint and cataract are major ocular disorders that majorly affect visual perception and interaction capability. This paper proposes a real-time video-b

researcharxiv-cs-cv
9 Jul 2026
Agents

A Three-Layer Framework for AI in Scientific Discovery

DGX agent

arXiv:2606.13566v2 Announce Type: replace Abstract: Current discussions of AI in scientific discovery are often dominated by two visible capabilities: search over existing knowledge and execution thro

agentsarxiv-cs-ai
8 Jul 2026
Safety

Agentic AI for Commercial Insurance Underwriting with Adversarial Self-Critique

DGX agent

arXiv:2602.13213v2 Announce Type: replace Abstract: Commercial insurance underwriting is a labor-intensive process that requires manual review of extensive documentation to assess risk and determine p

safetyarxiv-cs-ai
8 Jul 2026
Model Releases

Auto-DSM Under the Lens: A Black-Box Evaluation Framework for LLM-Based DSM Generation

DGX agent

arXiv:2607.05985v1 Announce Type: new Abstract: This paper presents a black-box evaluation framework to systematically assess the ability of Large Language Models (LLMs) to generate Design Structure M

model-releasesarxiv-cs-ai
8 Jul 2026
Industry

Benchmarking Coding Agents on Databricks’ Multi-Million Line Codebase

DGX agent

This Databricks blog post evaluates the performance and capabilities of coding agents when applied to real-world scenarios involving their own multi-million line codebase, likely assessing metrics suc

industrydatabricks
8 Jul 2026
Applications

Beyond Accuracy: How Humans Evaluate Legally Correct but Socially Controversial Legal Advice from Machines

DGX agent

arXiv:2607.05680v1 Announce Type: cross Abstract: AI systems are increasingly used to provide legal advice, raising questions about whether laypeople accept guidance from algorithms--especially when t

applicationsarxiv-cs-ai
8 Jul 2026
← Previous
1…6465666768…104
Next →