AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,983 results
31 Jul 2026

StealthBench: Measuring Operational Stealth in Autonomous Offensive-Security Agents

Model ReleasesDGX agent

arXiv:2607.26314v1 Announce Type: cross Abstract: Stealth, the discipline of achieving an objective without revealing your presence, capabilities, or collected intelligence, is what separates sophisti

Towards Generalized Synapse Detection Across Invertebrate Species

Model ReleasesDGX agent

arXiv:2509.17041v2 Announce Type: replace Abstract: Behavioural differences across organisms, whether healthy or pathological, are closely tied to the structure of their neural circuits. Yet, the fine

TraceCoder: Explainable and Auditable Code Generation with Position-Key Snippet Versioning

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.26307v1 Announce Type: new Abstract: Contemporary LLM-based coding agents produce code as black-box outputs: the rationale behind each line is hidden, the evolution of the code through benc

30 Jul 2026

A large-scale corpus of religious radio broadcast transcripts from webstream recordings in the United States

ResearchDGX agent

arXiv:2607.26249v1 Announce Type: new Abstract: Religious radio is a widespread but understudied form of mass communication in the United States, and content-level analysis of it has been constrained

AlloyDB adds group authentication to secure enterprise scale and AI agents

SafetyDGX agent

Database security traditionally relies on a fragile balance between the granular control developers need and the administrative overhead of managing thousands of individual database passwords. Between

EvoPINN: Agentic Discovery of Executable Algorithms for Physics-Informed Neural Networks

Model ReleasesDGX agent

arXiv:2607.26490v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) have emerged as a powerful paradigm for solving partial differential equations (PDEs), yet their performance

Gated Adaptation for Continual Learning in Human Activity Recognition

Model ReleasesDGX agent

arXiv:2603.10046v2 Announce Type: replace Abstract: Wearable sensors in Internet of Things (IoT) ecosystems increasingly support applications such as remote health monitoring, elderly care, and smart

Google DeepMind debuts Gemini Robotics 2 model series for humanoid robots

Model ReleasesDGX agent

Alphabet Inc.’s artificial intelligence research lab today debuted a family of models optimized to power humanoid robots. Google DeepMind says that the Gemini Robotics 2 series enables multiple autono

I'm actually fairly bearish on frontier lab valuations. I've never seen the reasons articulated to my satisfaction, so before I go to sleep,…

Model ReleasesDGX agent

I'm actually fairly bearish on frontier lab valuations. I've never seen the reasons articulated to my satisfaction, so before I go to sleep, I wanted to quickly jot down my thinking here. The basic is

Inference meta-monitoring for Amazon SageMaker AI endpoints with Amazon Quick

TutorialsDGX agent

Learn how to build an inference meta-monitoring system for Amazon SageMaker AI endpoints using Amazon Quick. This governance layer sits above production ML inference pipelines to continuously track pr

Investigating three real-world incidents in our cybersecurity evaluations

Model ReleasesDGX agent

Investigating three real-world incidents in our cybersecurity evaluations It happened again! This is turning into something of a pattern. Last week OpenAI accidentally exploited Hugging Face when one

MindForge: Teaching Small Language Models Whole-Life-Cycle Software Engineering via Source-Free Program Synthesis

AgentsDGX agent

arXiv:2607.27146v1 Announce Type: cross Abstract: Coding agents have made substantial progress on software engineering tasks that modify existing codebases, including bug fixing and feature implementa

Neural Architecture Search for Traffic Prediction: A Survey of Methods, Challenges, and Future Directions

ResearchDGX agent

arXiv:2607.26467v1 Announce Type: new Abstract: Traffic prediction is a core task in intelligent transportation systems, supporting applications such as adaptive signal control, route guidance, and ri

P.A.I. — Sleek Native Desktop AI Overlayer for Local Ollama Models 🤖⚡

Model ReleasesDGX agent

Greetings Community! 👋 I hope everyone is doing well! I'm Tauhid — Senior EEE student from a Bangladeshi University Today I'd like to share an open-source project I’ve been developing called P.A.I. (P

Robostreet Flow: A Lightweight, Ultra-Low-Drag Electric Tractor and Four-Truck Hybrid Convoy Architecture for Minimum-Cost Point-to-Point Freight

SafetyDGX agent

arXiv:2607.26250v1 Announce Type: new Abstract: Line-haul trucking costs are dominated by three comparably sized components: energy, driver labor, and equipment. Most efficiency technologies address o

SciFigQual-Bench: A Benchmark for Scientific Figure Quality Assessment with Full-Manuscript Context

Model ReleasesDGX agent

arXiv:2607.27084v1 Announce Type: new Abstract: Scientific images are the core elements of presenting experimental conclusions, elaborating system architecture, and supporting comparative arguments in

29 Jul 2026

AI-Assisted Knowledge Access for Legacy Enterprise Asset Management in Energy Operations: A Practical Retrieval System

TutorialsDGX agent

arXiv:2607.24792v1 Announce Type: cross Abstract: Energy utilities still run engineering work management, engineering procurement, and inventory processes on long-lived enterprise asset management pla

AI Security Leaderboard: benchmarking model robustness [P]

Model ReleasesDGX agent

We developed a leaderboard ranking frontier model security. There's no shortage of model capability rankings, but we didn't find anything comparable for model security. Yet security is becoming increa

AVE-Compass: Towards Holistic Evaluation for Audio-Video Editing Abilities

Model ReleasesDGX agent

arXiv:2607.24821v1 Announce Type: cross Abstract: While instruction-based video editing has advanced rapidly, real-world videos contain tightly coupled audio and visual signals, and editing one modali

Beyond Epistemia: Epistemic Schizologia and Large Language Models as Techno-Semiotic Machines

AgentsDGX agent

arXiv:2607.25620v1 Announce Type: new Abstract: Quattrociocchi and colleagues warn that the fluent outputs of large language models may allow linguistic plausibility to substitute for epistemic evalua

Building Large-Scale English-Romanian Literary Translation Resources with Open Models

Model ReleasesDGX agent

arXiv:2509.07829v4 Announce Type: replace-cross Abstract: Literary translation has recently gained attention as a distinct and complex task in machine translation research, yet translation by small op

CADENCE: A Cardiac Atom Dictionary for Interpretable Neural Concept Extraction from ECG Foundation Models

ResearchDGX agent

arXiv:2607.25244v1 Announce Type: new Abstract: Foundation models for 12-lead electrocardiograms (ECGs) transfer well across clinical tasks, but the physiological knowledge encoded in their representa

Chart-Supported or Model-Supplied? Examining MLLM-Generated Claims for Accessible Visualization

Model ReleasesDGX agent

arXiv:2607.25021v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) can connect visualization patterns to external causes, consequences, and domain knowledge, but the evidential b

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition

Model ReleasesDGX agent

arXiv:2607.25294v1 Announce Type: cross Abstract: Real-world tasks often require models to learn from task-specific context rather than relying only on pre-trained knowledge. While recent work has hig

Cyber-Capable AI Agents: Vulnerabilities, Evaluation Containment, and Defensive Response

AgentsDGX agent

arXiv:2607.25379v1 Announce Type: new Abstract: Cyber-capable AI agents combine language models with tools, memory, and execution en- vironments to perform multi-step offensive-security tasks. Existin

DocAnnot -- Accelerating the Creation of Key Information Extraction Datasets with GenAI-Powered Auto-annotation

ResearchDGX agent

arXiv:2607.24745v1 Announce Type: cross Abstract: Key Information Extraction (KIE) is vital for many document applications, but creating training datasets is traditionally a time-consuming manual proc

Does Runtime Topology Context Improve LLM-Generated Kubernetes Security Patches?

SafetyDGX agent

arXiv:2607.25995v1 Announce Type: cross Abstract: Kubernetes is central to the cloud-native ecosystem, orchestrating containerised workloads. Recent work suggests that large language models (LLMs) can

Grounded in Consensus, In Step With Emerging Science: A Consensus-Anchored Multi-Corpus Clinical Chatbot for Long COVID

ResearchDGX agent

arXiv:2607.25038v1 Announce Type: cross Abstract: Long COVID (LC) poses a challenge for clinical decision support because relevant evidence is distributed across sources with different update cycles,

HOLODECK 2.0: Vision-Language-Guided 3D World Generation with Editing

ResearchDGX agent

arXiv:2508.05899v3 Announce Type: replace Abstract: 3D scene generation plays a crucial role in gaming, artistic creation, virtual reality, and many other domains. However, current 3D scene design sti

i'll say it plainly, hermes agent desktop is the best agentic app i've used, and i'm a little mad i didn't find it sooner. it auto see the m…

Local AiDGX agent

i'll say it plainly, hermes agent desktop is the best agentic app i've used, and i'm a little mad i didn't find it sooner. it auto see the models i'm serving, laguna s 2.1 sitting on my dgx spark and

LivingArena: Do LLMs Know What Other LLMs Don't? Peer-Probing as Scalable Evaluation

ResearchDGX agent

arXiv:2607.24780v1 Announce Type: new Abstract: Evaluating frontier LLMs is challenging: static benchmarks suffer from contamination and saturation -- leaving users unable to distinguish top models an

LLM Scheming Inversely Scales with Pretraining Language Coverage

SafetyDGX agent

arXiv:2607.24769v1 Announce Type: new Abstract: With the growing capabilities of frontier models, AI alignment becomes increasingly critical in high-risk deployment settings. While recent work has emp

Lowering the implementation barrier of neutral-atom quantum computing with agentic workflows

AgentsDGX agent

arXiv:2607.25834v1 Announce Type: cross Abstract: Quantum computers are moving from research laboratories to industrial machines accessible via the cloud and integrated into high-performance computing

Measuring the State of Open Science in Transportation Using Large Language Models

ResearchDGX agent

arXiv:2601.14429v2 Announce Type: replace-cross Abstract: Open science initiatives have strengthened scientific integrity and accelerated research progress across many fields, but the state of their p

RefBench-PRO: Perceptual and Reasoning Oriented Benchmark for Referring Expression Comprehension

Model ReleasesDGX agent

arXiv:2512.06276v3 Announce Type: replace-cross Abstract: Referring Expression Comprehension (REC) is a vision-language task that localizes a specific image region based on a textual description. Exis

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement

Model ReleasesDGX agent

arXiv:2607.25886v1 Announce Type: cross Abstract: Recursive self-improvement requires turning evidence of model failures into better models. Data-centric post-training research entails diagnosing capa

SAFAARI: Schema-Aware Framework for Accelerated Advertiser Response Intelligence

AgentsDGX agent

arXiv:2607.25042v1 Announce Type: new Abstract: The evolution of customer support systems is rapidly advancing with agentic chatbots, yet these systems face significant limitations when accessing ente

ScalableRAG: High-Quality RAG at Zero Ingestion Cost

ResearchDGX agent

arXiv:2607.25135v1 Announce Type: new Abstract: Recent advances in RAG aim to optimize for performance by paying high ingestion costs for knowledge ingestion: building knowledge graphs or extracting S

Sense it with your eyes: Sensation Generation and Understanding for Advertisements

Model ReleasesDGX agent

arXiv:2607.25314v1 Announce Type: new Abstract: Sensory advertising evokes human senses through visual cues, enabling audiences to mentally simulate experiences and increasing persuasive impact. Despi

The borderless Lakehouse: Bring AWS, Databricks and Snowflake data to your AI agents

Model ReleasesDGX agent

Today’s data lakehouse is no longer mere data repository, but increasingly a system of action, actively executing tasks via always-on, autonomous AI agents. Rather than waiting for static reports, the

The Disruptive Impact of Large Language Models on Capture the Flag Competitions and the Path Toward Fair Play

SafetyDGX agent

arXiv:2607.25425v1 Announce Type: new Abstract: Capture the Flag (CTF) competitions are among cybersecurity's most effective training grounds, developing practical skill across cryptography, web explo

Verification Without Distrust: Reframing User-Side Oversight as Routine Epistemic Governance in Everyday Human-Chatbot Interaction

ResearchDGX agent

arXiv:2607.24761v1 Announce Type: cross Abstract: Research on human-AI interaction has long framed verification of system outputs as a trust-contingent behavior that better-calibrated trust should red

28 Jul 2026

A Frozen 12B Beats Frontier Models on Verified Work: 100% Accuracy, 0 Tokens, Bit-Exact, Forever

Model ReleasesDGX agent

arXiv:2607.23806v1 Announce Type: cross Abstract: Improving a language model today means retraining it: enormous compute, a new opaque model each cycle, non-deterministic output. We take the opposite

Accuracy potential of visual localization exploiting high-end street-level imagery

AgentsDGX agent

arXiv:2607.24409v1 Announce Type: new Abstract: Accurate and reliable pose information with respect to a reference frame is increasingly demanded across applications such as autonomous navigation, sur

AIR-BENCH Live: An Evolving Safety Benchmark for Foundation Models

Model ReleasesDGX agent

arXiv:2607.22671v1 Announce Type: new Abstract: Foundation-model safety benchmarks capture the AI risks of their time of publication: as models improve and governments pass new AI-safety legislation,

An Agentic Orchestration of Atomistic Simulations

AgentsDGX agent

arXiv:2607.22596v1 Announce Type: new Abstract: Atomistic simulations are central to materials design, but their execution involves complex, multi-step workflows that require significant human experti

An Interactive Vision Language Platform for Cognitive Remediation in Schizophrenia

ResearchDGX agent

arXiv:2607.22721v1 Announce Type: cross Abstract: Cognitive remediation tasks often require patients to perform structured actions involving object manipulation and sequential reasoning. For patients

Beyond Shapley: An Influence-Based Data Auditing Pipeline for LLM Alignment and Evaluation

Model ReleasesDGX agent

arXiv:2607.22766v1 Announce Type: cross Abstract: The alignment of Large Language Models (LLMs) is increasingly bottlenecked by data quality. As datasets scale, massive preference and instruction-tuni

Building AI That Works: ESnet's Pragmatic Approach to AI-Driven Operational Excellence

AgentsDGX agent

arXiv:2607.22948v1 Announce Type: cross Abstract: The ORBIT (Operations Responses and Business Intelligence Toolkit) project was initiated to assess agentic AI for the upcoming ESnet 7 initiative and

Can an Actor-Critic Optimization Framework Improve Analog Design?

ResearchDGX agent

arXiv:2603.24714v2 Announce Type: replace Abstract: Analog design often slows down because even small changes to device sizes or biases require expensive simulation cycles, and high-quality solutions

ConsistencyGate: Preventing Memory Contamination in LLM Agents via Self-Consistency Admission Control

Model ReleasesDGX agent

arXiv:2607.22962v1 Announce Type: new Abstract: LLM agents that operate over many turns accumulate facts in an external memory store and reuse them as premises for downstream reasoning. A hallucinated

Decentralized Granular Access Control for Agentic AI Systems in Critical Infrastructure

Model ReleasesDGX agent

arXiv:2607.22611v1 Announce Type: new Abstract: The deployment of autonomous AI agents in production infrastructure introduces fundamental security challenges that traditional role-based access contro

DRC-Aid: Design-Rule Correction via Agentic Framework utilizing Inference-Time Large Language Models

Local AiDGX agent

arXiv:2607.22761v1 Announce Type: cross Abstract: Resolving Design Rule Violations (DRVs) in layouts entails an iterative loop of geometric edits and verification. We present DRC-Aid, a closed-loop ag

DreamStyle3D: Efficient 3D Stylized Asset Generation via Dual-Attention Disentanglement

ResearchDGX agent

arXiv:2607.24721v1 Announce Type: new Abstract: With the growth of gaming, animation, and virtual reality industries, the demand for efficient generation of stylized 3D assets is rapidly increasing. H

DuoAD: Leveraging [CLS] Dual Characteristics for Training-Free Few-Shot Anomaly Detection

Model ReleasesDGX agent

arXiv:2607.23924v1 Announce Type: cross Abstract: Vision foundation models have enabled strong training-free anomaly detection (AD). However, most existing approaches rely primarily on independent loc

Energy Constrained Hierarchical Underwater Monitoring via Local Multi-Agent RAG

HardwareDGX agent

arXiv:2607.24313v1 Announce Type: cross Abstract: Marine life monitoring is limited by strict energy constraints, poor underwater connectivity, and the high cost of transmitting raw multimodal data fr

EviBack: Search-Agent Reinforcement Learning via Evidence-Constrained Teacher Backoff

AgentsDGX agent

arXiv:2607.23955v1 Announce Type: new Abstract: Reinforcement learning enables Agentic RAG systems to learn multi-turn search from verifiable outcome rewards, but all- zero rollout groups provide no c

Failures Reveal What Metrics Miss: An Evidence-Driven Agent for Recursive Refinement of ECG Classifiers

AgentsDGX agent

arXiv:2607.24419v1 Announce Type: new Abstract: Deep models have substantially advanced 12-lead ECG classification, yet their refinement still relies heavily on human experts to inspect failures and i

GazeLT: Visual attention-guided long-tailed disease classification in chest radiographs

TutorialsDGX agent

arXiv:2508.09478v2 Announce Type: replace Abstract: In this work, we present GazeLT, a human visual attention integration-disintegration approach for long-tailed disease classification. A radiologist'

Guiding Language Models to Be More Empathetic: Culturally Sensitive Mental Health Advice Generation Through Human-LLM Collaboration

Model ReleasesDGX agent

arXiv:2607.23538v1 Announce Type: new Abstract: Despite recent advances in large language models (LLMs), their ability to generate empathetic mental health counseling responses in low-resource languag

← Previous
1…4849505152…84
Next →