AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,980 results
Tutorials

Automatic Layer Selection for Hallucination Detection

DGX agent

arXiv:2605.26366v1 Announce Type: new Abstract: Recent studies on hallucination detection have shown that hallucination-related signals are more strongly encoded in intermediate layers than in the fin

tutorialsarxiv-cs-ai
27 May 2026
Applications
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Been using Grok Build these past few days, and the thing that really got me hooked is Imagine and Imagine Video. I built a full dinosaur enc…

DGX agent

Been using Grok Build these past few days, and the thing that really got me hooked is Imagine and Imagine Video. I built a full dinosaur encyclopedia site — every image, every video clip on it, all ge

applicationselon-musk--x
27 May 2026
Research

Beyond Binary: Speech Representations Across the Cognitive Score Hierarchy

DGX agent

arXiv:2605.27189v1 Announce Type: new Abstract: This study examines the relationship between speech representations and the hierarchical structure of cognitive assessment in mild cognitive impairment.

researcharxiv-cs-cl
27 May 2026
Model Releases

Beyond Holistic Models: Systematic Component-level Benchmarking of Deep Multivariate Time-Series Forecasting

DGX agent

arXiv:2605.26562v1 Announce Type: new Abstract: While previous research in multivariate time series forecasting has focused on developing complex holistic models, this work advocates for a shift towar

model-releasesarxiv-cs-lg
27 May 2026
Agents

Building self-improving tax agents with Codex

DGX agent

This article describes how OpenAI's Codex model can be used to build autonomous tax agents capable of self-improvement through code generation and execution. The work demonstrates using large language

agentsopenai
27 May 2026
Model Releases

CNNs, Transformers, Hybrid, and Vision Language Models for Skin Cancer Detection

DGX agent

arXiv:2605.26294v1 Announce Type: new Abstract: Skin cancer is a common and fast rising malignancy worldwide. Early detection is critical for improving outcomes. Deep learning models trained on dermos

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Drive-P2D: A Progressive Perception-to-Decision Benchmark for VLMs in Autonomous Driving

DGX agent

arXiv:2601.14702v2 Announce Type: replace Abstract: Autonomous driving requires reliable perception and safe decision-making in complex scenarios. Recent vision-language models (VLMs) demonstrate reas

model-releasesarxiv-cs-ai
27 May 2026
Local Ai

Experiments in Agentic AI for Science

DGX agent

arXiv:2605.26305v1 Announce Type: new Abstract: This paper details two novel frameworks for developing autonomous, agentic AI in scientific workflows. Both systems leverage a hybrid Local Body, Remote

local-aiarxiv-cs-ai
27 May 2026
Safety

From Norms to Indicators (N2I-RAG): An Agentic Retrieval-Augmented Generation Framework for Legal Indicator Computation

DGX agent

arXiv:2605.26926v1 Announce Type: new Abstract: Computing legal indicators from normative texts is a key task in legal monitoring and policy evaluation, but presents significant challenges due to the

safetyarxiv-cs-ai
27 May 2026
Model Releases

From PDF to RAG-Ready: Evaluating Document Conversion Frameworks for Domain-Specific Question Answering

DGX agent

arXiv:2604.04948v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems depend critically on the quality of document preprocessing, yet no prior study has evaluated PDF

model-releasesarxiv-cs-ai
27 May 2026
Research

Gumbel Machine: Counterfactual Student Writing Generation via Gumbel Noise Steering

DGX agent

arXiv:2605.27249v1 Announce Type: new Abstract: An effective method of teaching across disciplines is to provide examples of high-quality work. However, an example may be significantly different from

researcharxiv-cs-ai
27 May 2026
Local Ai

HRVConformer: Neonatal Hypoxic-Ischemic Encephalopathy Classification from the Heart Rate signals

DGX agent

arXiv:2605.26190v1 Announce Type: cross Abstract: This paper presents the HRVConformer, a novel deep learning architecture for the classification of hypoxic-ischemic encephalopathy (HIE) using the ins

local-aiarxiv-cs-ai
27 May 2026
Agents

I really appreciate the lessons and technical ideas @samaysham & team were able to share about their tax agent system, which learns from pro…

DGX agent

I really appreciate the lessons and technical ideas @samaysham & team were able to share about their tax agent system, which learns from production traces to self-improve via detailed tracing tightly

agentslinus-lee--x
27 May 2026
Model Releases

I think Anthropic and OpenAI have found product-market fit

DGX agent

Anthropic are strongly rumored to be about to have their first profitable quarter. Stories are circulating of companies surprised at how expensive their LLM bills are becoming from usage by their staf

model-releasessimon-willison
27 May 2026
Safety

LAD-VF: LLM-Automatic Differentiation Enables Fine-Tuning-Free Robot Planning from Formal Methods Feedback

DGX agent

arXiv:2509.18384v2 Announce Type: replace Abstract: Large language models (LLMs) can translate natural language instructions into executable action plans for robotics, autonomous driving, and other do

safetyarxiv-cs-ro
27 May 2026
Model Releases

LiveK12Bench: Have Large Multimodal Models Truly Conquered High School-level Examinations?

DGX agent

arXiv:2605.26781v1 Announce Type: new Abstract: Advanced Large Multimodal Models (LMMs) have demonstrated impressive performance in K-12 reasoning tasks, exhibiting great promise as intelligent tutors

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

LURE: Live-Usage Replay Evaluations for Reducing Evaluation Awareness

DGX agent

arXiv:2605.26438v1 Announce Type: cross Abstract: Large language models can recognize when they are being evaluated (evaluation awareness) and behave differently because of that, which undermines the

model-releasesarxiv-cs-ai
27 May 2026
Tutorials

Optimising Factual Consistency in Summarisation via Preference Learning from Multiple Imperfect Metrics

DGX agent

arXiv:2605.26840v1 Announce Type: new Abstract: Reinforcement learning with evaluation metrics as rewards is widely used to enhance specific capabilities of language models. However, for tasks such as

tutorialsarxiv-cs-cl
27 May 2026
Local Ai

Prototyping an End-to-End Multi-Modal Tiny-CNN for Cardiovascular Sensor Patches

DGX agent

arXiv:2510.18668v2 Announce Type: replace-cross Abstract: The vast majority of cardiovascular diseases may be preventable if early signs and risk factors are detected. Cardiovascular monitoring with b

local-aiarxiv-cs-cv
27 May 2026
Local Ai

ReVEL: Multi-Turn Reflective LLM-Guided Heuristic Evolution via Structured Performance Feedback

DGX agent

arXiv:2604.04940v2 Announce Type: replace Abstract: Designing effective heuristics for NP-hard combinatorial optimization problems remains challenging and often requires substantial domain expertise.

local-aiarxiv-cs-ai
27 May 2026
Model Releases

RoadGIE: Towards A Global-Scale Aerial Benchmark for Generalizable Interactive Road Extraction

DGX agent

arXiv:2605.26862v1 Announce Type: new Abstract: Accurate road segmentation from aerial imagery is fundamental to many geospatial applications. However, existing datasets often suffer from limited scen

model-releasesarxiv-cs-cv
27 May 2026
Applications

Role-Based Access Control for Humans and Agents

DGX agent

This article discusses implementing role-based access control (RBAC) systems that work for both human users and AI agents, likely addressing how to manage permissions and authentication in environment

applicationsmodal-blog
27 May 2026
Model Releases

SpaceVista: All-Scale Visual Spatial Reasoning from mm to km

DGX agent

arXiv:2510.09606v2 Announce Type: replace Abstract: With the current surge in spatial reasoning explorations, researchers have made significant progress in understanding indoor scenes, but still strug

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

SteelDS: A High-Resolution Video Dataset of E40 Steel Scrap for Object Detection and Instance Segmentation

DGX agent

arXiv:2605.26682v1 Announce Type: cross Abstract: This dataset provides high-resolution, annotated video sequences of shredded E40-grade steel and copper scrap on a conveyor belt. Captured in a contro

model-releasesarxiv-cs-cv
27 May 2026
Research

Structure-Adaptive Conformal Inference for Large-Scale Out-of-Distribution Testing

DGX agent

arXiv:2605.26429v1 Announce Type: cross Abstract: This paper addresses structured out-of-distribution (OOD) testing in high-stakes machine learning applications. Traditional conformal methods rely on

researcharxiv-cs-ai
27 May 2026
Applications

Towards Real-World Identification of Fatigued Muscle Groups via Musculoskeletal Simulation

DGX agent

arXiv:2605.26151v1 Announce Type: cross Abstract: Contactless diagnosis of musculoskeletal disorders can potentially improve population health as well as robot behaviours in collaborative settings. Ho

applicationsarxiv-cs-ro
27 May 2026
Applications

UltraCUA: A Foundation Model for Computer Use Agents with Hybrid Action

DGX agent

arXiv:2510.17790v3 Announce Type: replace-cross Abstract: Computer-use agents face a fundamental limitation. They rely exclusively on primitive GUI actions (click, type, scroll), creating brittle exec

applicationsarxiv-cs-cl
27 May 2026
Research

Understanding the Challenges in Iterative Generative Optimization with LLMs

DGX agent

arXiv:2603.23994v2 Announce Type: replace-cross Abstract: Generative optimization uses large language models (LLMs) to iteratively improve artifacts (such as code, workflows or prompts) using executio

researcharxiv-cs-ai
27 May 2026
Hardware

Xe-Forge: Multi-Stage LLM-Powered Kernel Optimization for Intel GPU

DGX agent

arXiv:2605.26118v1 Announce Type: cross Abstract: Porting deep learning algorithms to new hardware accelerators requires developers to repeatedly apply the same low-level optimizations -- quantization

hardwarearxiv-cs-ai
27 May 2026
Model Releases

Zero-Shot MARL Benchmark in the Cyber-Physical Mobility Lab

DGX agent

arXiv:2601.16578v2 Announce Type: replace Abstract: We present a reproducible benchmark for evaluating sim-to-real transfer of Multi-Agent Reinforcement Learning (MARL) policies for Connected and Auto

model-releasesarxiv-cs-ro
27 May 2026
Agents

A Multi-Agent LLM Framework for Rating the Quality of Surgical Feedback

DGX agent

arXiv:2605.25440v1 Announce Type: cross Abstract: Verbal feedback delivered by attending surgeons in the operating room plays a critical formative role in resident trainee skill acquisition. Yet, asse

agentsarxiv-cs-ai
26 May 2026
Agents

AgentWatch: Proactive AWS monitoring with ambient agents

DGX agent

In this post, we demonstrate the capabilities of AgentWatch through practical implementation. You will see how the solution performs infrastructure checks every 15 minutes, summarizing CloudWatch metr

agentsaws-ml-blog
26 May 2026
Agents

Architecting Agentic Communities using Design Patterns

DGX agent

arXiv:2601.03624v3 Announce Type: replace Abstract: The rapid evolution of Large Language Models (LLM) and subsequent Agentic AI technologies requires systematic architectural guidance for building so

agentsarxiv-cs-ai
26 May 2026
Safety

Auditing medical multi-agent AI reveals risks of false consensus

DGX agent

arXiv:2510.10185v2 Announce Type: replace-cross Abstract: Large language models are increasingly being assembled into medical multi-agent systems that emulate multidisciplinary consultation through sp

safetyarxiv-cs-ai
26 May 2026
Research

Beyond Control-Flow: Integrating the Resource Perspective into Multi-Collaborative Process Modeling from Text

DGX agent

arXiv:2605.24546v1 Announce Type: new Abstract: Process modeling is a sub-domain of Business Process Management (BPM) focused on the translation of process artifacts into formal models. This task trad

researcharxiv-cs-ai
26 May 2026
Model Releases

Beyond Final Answers: Auditing Trajectory-Level Hallucinations in Multi-Agent Industrial Workflows

DGX agent

arXiv:2605.24219v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as autonomous agents that reason, use tools, and act over multiple steps. Yet most hallucination

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

BODHI: Precise OS Kernel Specification Inference

DGX agent

arXiv:2605.23931v1 Announce Type: new Abstract: The formal verification of operating system kernels requires precise specifications that capture the intended behavior of system calls. Writing these sp

model-releasesarxiv-cs-ai
26 May 2026
Hardware

Build high-performance generative AI systems with Strands Agents, NVIDIA NIM, and Amazon Bedrock AgentCore

DGX agent

In this post you'll learn how to build a multi-agent campaign review system that demonstrates parallel reasoning, context persistence, and traceable execution paths using an integrated architecture th

hardwareaws-ml-blog
26 May 2026
Applications

Choosing to Stay Human

DGX agent

This essay by Ethan Mollick explores how individuals can maintain their humanity and agency in an increasingly AI-driven world, likely addressing practical strategies for preserving human skills, crea

applicationsethan-mollick
26 May 2026
Model Releases

Claw-Anything: Benchmarking Always-On Personal Assistants with Broader Access to User's Digital World

DGX agent

arXiv:2605.26086v1 Announce Type: new Abstract: Large language model agents are increasingly envisioned as always-on personal assistants with access to anything relevant in the user's digital world. Y

model-releasesarxiv-cs-ai
26 May 2026
Agents

CRISP -- Clustering-Based Redundancy-Reduced Instance Sampling for Pathology Case Representation and Retrieval

DGX agent

arXiv:2605.24253v1 Announce Type: cross Abstract: Digital pathology archives increasingly contain multiple whole-slide images (WSIs) per case, capturing spatially distinct tumour regions and reflectin

agentsarxiv-cs-ai
26 May 2026
Model Releases

DRInQ: Evaluating Conversational Implicature with Controlled Context Variation

DGX agent

arXiv:2605.24267v1 Announce Type: new Abstract: Human conversation relies heavily on conversational implicature, in which speakers convey meanings that are suggested rather than explicitly stated. Alt

model-releasesarxiv-cs-cl
26 May 2026
Safety

Dynamic Optimization and Safety Indicator Injection for Jailbreaking Text-to-Image Models with Multimodal Safety Filters

DGX agent

arXiv:2505.18979v2 Announce Type: replace Abstract: Text-to-image (T2I) models can generate not-safe-for-work (NSFW) content, motivating multi-stage safety pipelines with both text and image filters.

safetyarxiv-cs-lg
26 May 2026
Model Releases

Empirical Analysis and Detection of Hallucinations in LLM-Generated Bug Report Summaries

DGX agent

arXiv:2605.24137v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to generate summaries of software bug reports, including sections such as Steps-to-Reproduce (S2R),

model-releasesarxiv-cs-ai
26 May 2026
Safety

EPPC-OASIS: Ontology-Aware Adaptation and Structured Inference Refinement for Electronic Patient-Provider Communication Mining in Secure Messages

DGX agent

arXiv:2605.24172v1 Announce Type: new Abstract: Secure patient-provider messages contain clinically important communication behaviors that are difficult to characterize manually at scale. The Electron

safetyarxiv-cs-ai
26 May 2026
Safety

FairJudge: Abstention-Aware Multimodal Judges for Fairness and Alignment Evaluation in Text-to-Image Models

DGX agent

arXiv:2510.22827v3 Announce Type: replace-cross Abstract: Evaluating text-to-image (T2I) systems requires judging not only whether an image matches a prompt, but also whether socially salient attribut

safetyarxiv-cs-lg
26 May 2026
Model Releases

Faithfulness Metrics Don't Measure Faithfulness: A Meta-Evaluation with Ground Truth

DGX agent

arXiv:2605.25052v1 Announce Type: new Abstract: Chains of thought (CoTs) have become central in interpreting and auditing behaviors of large language models. Yet growing evidence suggests that these t

model-releasesarxiv-cs-cl
26 May 2026
Safety

GeoSVG-RL: Geometry-Aware Reinforcement Learning for Layout-Constrained Text-to-SVG Diagram Generation

DGX agent

arXiv:2605.25447v1 Announce Type: new Abstract: Generating structured, editable diagrams remains a significant challenge for contemporary large language models, despite their proficiency in general-pu

safetyarxiv-cs-cl
26 May 2026
← Previous
1…7980818283…104
Next →