AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
3,914 results
Model Releases

VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models

DGX agent

arXiv:2512.22539v2 Announce Type: replace-cross Abstract: While Vision-Language-Action models (VLAs) are rapidly advancing towards generalist robot policies, it remains difficult to quantitatively und

model-releasesarxiv-cs-cv
3 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

3DCodeBench: Benchmarking Agentic Procedural 3D Modeling Via Code

DGX agent

arXiv:2606.01057v1 Announce Type: cross Abstract: Procedural 3D modeling through code is emerging as a versatile paradigm, offering deterministic, engine-ready, and precisely editable assets that neur

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

An Asynchronous Two-Speed Kalman Filter for Real-Time UUV Cooperative Navigation Under Acoustic Delays

DGX agent

arXiv:2604.02878v2 Announce Type: replace Abstract: In Global Navigation Satellite System (GNSS)-denied underwater environments, individual unmanned underwater vehicles (UUVs) suffer from unbounded de

agentsarxiv-cs-ro
2 Jun 2026
Tutorials

An explainable hierarchical self attention-based approach for tremor detection in the time domain

DGX agent

arXiv:2606.00461v1 Announce Type: new Abstract: Tremor is a common movement disorder associated with conditions like Parkinson's disease and Essential tremor, traditionally diagnosed through expert cl

tutorialsarxiv-cs-cv
2 Jun 2026
Model Releases

An Open-Source Benchmark and Baseline for Multi-temporal Referring Segmentation

DGX agent

arXiv:2606.00987v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have shown strong visual understanding and language-guided grounding abilities, yet their capacity for multi-temp

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ASE-26: a curriculum for agentic software engineering as a discipline

DGX agent

arXiv:2606.01152v1 Announce Type: cross Abstract: The work of a professional software engineer has begun to consist, increasingly, of directing agents rather than writing code, and the empirical evide

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Attention mechanisms and transfer learning for robust peach leaf damage classification under domain shift

DGX agent

arXiv:2606.02045v1 Announce Type: cross Abstract: Artificial intelligence provides a practical framework for crop damage assessment from imagery data, supporting early decision-making in agricultural

model-releasesarxiv-cs-ai
2 Jun 2026
Applications

AutoForest: Automatically Generating Forest Plots from Biomedical Studies with End-to-End Evidence Extraction and Synthesis

DGX agent

arXiv:2606.02403v1 Announce Type: cross Abstract: Systematic reviews rely on forest plots to synthesise quantitative evidence across biomedical studies, but generating them remains a fragmented and la

applicationsarxiv-cs-ai
2 Jun 2026
Local Ai

AutoIQ: An Ensemble Framework for Automatic Assessment of Geometric Distortion in Prostate Diffusion-Weighted Imaging

DGX agent

arXiv:2606.00393v1 Announce Type: cross Abstract: Geometric distortion in prostate diffusion-weighted imaging (DWI) can impair lesion localization and reduce the reliability of MRI-based clinical asse

local-aiarxiv-cs-cv
2 Jun 2026
Safety

Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation

DGX agent

arXiv:2602.11790v2 Announce Type: replace Abstract: Although recent end-to-end video generation models demonstrate impressive performance in visually oriented content creation, they remain limited in

safetyarxiv-cs-ai
2 Jun 2026
Agents

BlueME: Robust Underwater Robot-to-Robot Communication Using Compact Magnetoelectric Antennas

DGX agent

arXiv:2411.09241v5 Announce Type: replace Abstract: We present the design, development, and experimental validation of BlueME, a compact magnetoelectric (ME) antenna array system for underwater robot-

agentsarxiv-cs-ro
2 Jun 2026
Research

Bridging Topology and Deep Representation Learning: A TDA-ViT Fusion Model for Four-Class Brain Tumor Classification

DGX agent

arXiv:2606.00927v1 Announce Type: new Abstract: Accurate brain tumor classification from magnetic resonance imaging (MRI) is a key requirement for early diagnosis and clinical decision-making. Vision

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures

DGX agent

arXiv:2505.24069v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are deployed on increasingly complex tasks that require multi-step decision-making. Understanding their algorithm

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Citation Grounding: Detecting and Reducing LLM Citation Hallucinations via Legal Citation Graphs

DGX agent

arXiv:2606.00898v1 Announce Type: new Abstract: Large language models systematically hallucinate legal citations -- fabricating statute references, citing repealed provisions, and confusing jurisdicti

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree

DGX agent

arXiv:2606.01494v1 Announce Type: cross Abstract: Agent skills extend AI agents with reusable instructions, tools, scripts, references, and workflows, establishing a security boundary distinct from bo

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Constitutional Black-Box Monitoring for Scheming in LLM Agents

DGX agent

arXiv:2603.00829v2 Announce Type: replace-cross Abstract: Safe deployment of Large Language Model (LLM) agents in autonomous settings requires reliable oversight mechanisms. A central challenge is det

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

Context Matters: Repository-Aware Security Analysis of the Agent Skill Ecosystem

DGX agent

arXiv:2603.16572v2 Announce Type: replace-cross Abstract: Agent skills extend local AI agents, such as Claude Code and OpenClaw, with additional functionality. Their growing popularity has led to dedi

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

CountGD++: Generalized Prompting for Open-World Counting

DGX agent

arXiv:2512.23351v2 Announce Type: replace Abstract: The flexibility and accuracy of methods for automatically counting objects in images and videos are limited by the way the object can be specified.

agentsarxiv-cs-cv
2 Jun 2026
Model Releases

Cross-Generational Transfer of Adversarial Attacks Reveals Non-Monotonic Safety Alignment in LLMs

DGX agent

arXiv:2606.00813v1 Announce Type: cross Abstract: Safety alignment in LLMs does not improve monotonically across model generations. Studying four generations of Google's Gemma family (7B-31B) with qua

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

DetailMaster: Can Your Text-to-Image Model Handle Long Prompts?

DGX agent

arXiv:2505.16915v3 Announce Type: replace-cross Abstract: While recent Text-to-Image (T2I) models show impressive capabilities in synthesizing images from brief descriptions, they struggle with the lo

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Early Prediction of Liver Cirrhosis Up to Two Years in Advance: A Machine Learning Study Benchmarking Against the FIB-4 and APRI Scores

DGX agent

arXiv:2601.00175v2 Announce Type: replace Abstract: Objective: Develop and evaluate machine learning (ML) models for predicting incident liver cirrhosis (LC) one and two years prior to diagnosis using

model-releasesarxiv-cs-lg
2 Jun 2026
Research

Edge-Based QoS-Aware Adaptive Task Placement: A Closed-Loop Control in Multi-Robot Systems

DGX agent

arXiv:2606.00552v1 Announce Type: cross Abstract: Multi-robot systems (MRS) increasingly offload compute-intensive perception tasks to edge nodes to meet strict time-sensitive Quality-of-Service (QoS)

researcharxiv-cs-ro
2 Jun 2026
Model Releases

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning

DGX agent

arXiv:2604.03893v2 Announce Type: replace Abstract: Current multimodal benchmarks for scientific reasoning primarily evaluate local information extraction -- models recognize symbols and values and th

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

FigSIM: A Dataset for Fine-grained Suicide Severity and Figurative Language in Suicide Memes

DGX agent

arXiv:2606.02523v1 Announce Type: new Abstract: Suicide memes are memes used to express suicide-related thoughts or comment on suicide-related issues. Suicide memes are increasingly common on social m

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

FVSpec: Real-World Property-Based Tests as Lean Challenges

DGX agent

arXiv:2606.01008v1 Announce Type: cross Abstract: We present a benchmark for evaluating AI models and agents on real-world formal software verification tasks. We first scrape 11,039 property-based tes

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Hierarchical Online Prompt Mutation with Dual-Loop Feedback for Guardrailed Evidence Document Generation: A Production-Evaluation Case Study

DGX agent

arXiv:2606.01472v1 Announce Type: cross Abstract: High-stakes production document-generation systems require language models to be adaptive, evidence-grounded, and auditable. We present HOPM, a hierar

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

HLL: Can Agents Cross Humanity's Last Line of Verification?

DGX agent

arXiv:2606.02449v1 Announce Type: new Abstract: Multimodal agents are increasingly expected to operate interfaces on behalf of users, raising a central deployment question: can they truly substitute f

model-releasesarxiv-cs-ai
2 Jun 2026
Tutorials

Human in the Loop Adaptive Optimization for Improved Time Series Forecasting

DGX agent

arXiv:2505.15354v2 Announce Type: replace Abstract: Time series forecasting models often produce systematic, predictable errors even in critical domains such as energy, finance, and healthcare. We int

tutorialsarxiv-cs-lg
2 Jun 2026
Safety

Isolating LLM Lexical Bias: A Curation-Free Triangulated Metric for Preference-Stage Learning

DGX agent

arXiv:2606.00334v1 Announce Type: cross Abstract: Various language domains have undergone remarkable changes in recent years; these shifts are largely attributed to the advent of Large Language Models

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

JenBridge: Adaptive Long-Form Video Soundtracking across Scene Transitions

DGX agent

arXiv:2606.01703v1 Announce Type: cross Abstract: We address the challenge of generating high-fidelity, long-form soundtracks that remain coherent across scene transitions. Existing AI music systems a

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

LLM Consortium for Software Design Refinement: A Controlled Experiment on Multi-Agent Collaboration Topologies

DGX agent

arXiv:2606.01490v1 Announce Type: cross Abstract: We present a controlled experiment evaluating 12 multi-agent LLM collaboration topologies for software architecture design. Using a 2imes2imes2 factor

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Monitoring Agentic Systems Before They're Reliable

DGX agent

arXiv:2606.02494v1 Announce Type: cross Abstract: Agentic systems entering production typically operate as partially integrated assemblies where structural defects, not task-level errors, dominate the

agentsarxiv-cs-ai
2 Jun 2026
Research

Not What, But How: A Communicative Audit of LLM Response Framing

DGX agent

arXiv:2606.02493v1 Announce Type: new Abstract: Large language models (LLMs) are being increasingly used to answer subjective, information-seeking questions, where users are sensitive to how responses

researcharxiv-cs-cl
2 Jun 2026
Model Releases

Order within Chaos: Capturing Intrinsic Energy Anomalies for AI-Manipulated Image Forgery Localization

DGX agent

arXiv:2606.02178v1 Announce Type: cross Abstract: Recent advancements in generative AI have led to image editing models capable of producing realistic forgeries that evade traditional image forgery lo

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Peacemaker at ATE-IT: Automatic term extraction from Italian text for waste management data using encoder model

DGX agent

arXiv:2606.01469v1 Announce Type: new Abstract: The development of automatic term extraction has become increasingly important in modern technology. Automatic term extraction can be found in virtually

researcharxiv-cs-cl
2 Jun 2026
Safety

Perspective on Bias in Biomedical AI: Preventing Downstream Healthcare Disparities

DGX agent

arXiv:2604.14514v2 Announce Type: replace Abstract: Healthcare disparities persist across socioeconomic boundaries, often attributed to unequal access to screening, diagnostics, and therapeutics. Howe

safetyarxiv-cs-ai
2 Jun 2026
Agents

RDA: Reward Design Agent for Reinforcement Learning

DGX agent

arXiv:2606.01672v1 Announce Type: new Abstract: Reinforcement learning has enabled the acquisition of impressive robotic skills, but typically requires hand-crafted reward functions that are slow to d

agentsarxiv-cs-lg
2 Jun 2026
Research

Rethinking Evaluation Paradigms in IBP-based Certified Training

DGX agent

arXiv:2606.02134v1 Announce Type: cross Abstract: Deep neural networks achieve strong performance on many supervised learning tasks but remain vulnerable to adversarial perturbations. Neural network v

researcharxiv-cs-ai
2 Jun 2026
Safety

RL-ACRGNet: Reinforcement Learning-Based Chest Radiology Report Generation Network

DGX agent

arXiv:2606.02035v1 Announce Type: new Abstract: Medical imaging interpretation is a foundational pillar of modern clinical diagnostics, yet the manual generation of radiology reports remains a time-co

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

RoboBenchMart: Benchmarking Robots in Retail Environment

DGX agent

arXiv:2511.10276v2 Announce Type: replace-cross Abstract: Most existing robotic manipulation benchmarks focus on tabletop or household scenarios. While these setups have driven impressive progress, it

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

RocketSmith: An Agentic System for High-Powered Rocket Design and Manufacturing

DGX agent

arXiv:2606.00097v1 Announce Type: new Abstract: This work presents RocketSmith, an agentic system capable of the design, manufacturing, and optimization processes in high powered rocket development. T

agentsarxiv-cs-ro
2 Jun 2026
Model Releases

Ryze: Evidence-Enriched Data Synthesis from Biomedical Papers

DGX agent

arXiv:2606.00902v1 Announce Type: new Abstract: General-purpose VLMs remain unreliable for biomedical research because valid answers in scientific papers depend on evidence split across figures, table

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Scaling Agentic Capabilities via Grounded Interaction Synthesis

DGX agent

arXiv:2606.02001v1 Announce Type: new Abstract: General agentic intelligence hinges on the ability to interact with diverse real-world tools to complete complex tasks, a capability fundamentally tied

agentsarxiv-cs-cl
2 Jun 2026
Local Ai

scicode-lint: Detecting Methodology Bugs in Scientific Python Code with LLM-Generated Patterns

DGX agent

arXiv:2603.17893v2 Announce Type: replace-cross Abstract: Methodology bugs in scientific Python code produce plausible but incorrect results that traditional linters and static analysis tools cannot d

local-aiarxiv-cs-ai
2 Jun 2026
Safety

SHERLOCK: Towards Dynamic Knowledge Adaptation in LLM-enhanced E-commerce Risk Management

DGX agent

arXiv:2510.08948v4 Announce Type: replace-cross Abstract: Effective e-commerce risk management requires in-depth case investigations to identify emerging fraud patterns in highly adversarial environme

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

SMH-Bench: Benchmarking LLM Agents for Environment-Grounded Reasoning and Action in Smart Homes

DGX agent

arXiv:2606.01912v1 Announce Type: new Abstract: Smart homes are evolving toward complex state-dependent living environments, requiring Large Language Models (LLMs) to reason over user intent, preferen

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

SortingHat: Redefining Operating Systems Education with a Tailored Digital Teaching Assistant

DGX agent

arXiv:2606.00015v1 Announce Type: cross Abstract: Operating Systems (OS) courses are among the most challenging in computer science education due to the complexity of internal structures and the diver

agentsarxiv-cs-ai
2 Jun 2026
Agents

SWE-rebench V2: Language-Agnostic SWE Task Collection at Scale

DGX agent

arXiv:2602.23866v2 Announce Type: replace-cross Abstract: Software engineering agents (SWE) are improving rapidly, with recent gains largely driven by reinforcement learning (RL). However, RL training

agentsarxiv-cs-cl
2 Jun 2026
← Previous
1…5960616263…82
Next →