AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “engineering”

GridTimelineEvolution
3,145 results
Agents

3100 Opinions on Code Review in an AI World: Building Causal Theory from Practitioner Discourse

DGX agent

arXiv:2607.07980v1 Announce Type: cross Abstract: Coding agents now author entire pull requests, and practitioners sharply disagree about what this does to code review: whether it becomes the bottlene

agentsarxiv-cs-ai
10 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Context Graphs for Proactive Enterprise Agents

DGX agent

arXiv:2607.07721v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) and agentic frameworks have advanced enterprise AI considerably, yet agents remain fundamentally reactive: they wai

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Infinity-Parser2 Technical Report

DGX agent

arXiv:2607.07836v1 Announce Type: new Abstract: We present Infinity-Parser2, a large multimodal model that couples a controllable data-synthesis pipeline with multi-task reinforcement learning for end

model-releasesarxiv-cs-ai
10 Jul 2026
Agents

Power and Limitations of Aggregation in Compound AI Systems

DGX agent

arXiv:2602.21556v2 Announce Type: replace Abstract: When designing compound AI systems, a common approach is to query multiple copies of the same model and aggregate the responses to produce a synthes

agentsarxiv-cs-ai
9 Jul 2026
Safety

KAT-Coder-V2.5 Technical Report

DGX agent

arXiv:2607.05471v1 Announce Type: cross Abstract: We present KAT-Coder-V2.5, a coding-focused agentic model trained to act autonomously inside real, executable repositories rather than as a single-tur

safetyarxiv-cs-ai
8 Jul 2026
Model Releases

OrchardBench: A Physically-Grounded, GPU-Parallel Apple-Orchard Simulation Benchmark for Agricultural Robotics

DGX agent

arXiv:2607.06337v1 Announce Type: cross Abstract: Robotic tree-fruit harvesting is a flagship problem for agricultural automation, but progress is bottlenecked by the cost and irreproducibility of fie

model-releasesarxiv-cs-cv
8 Jul 2026
Safety

Adaptive Inference Batching using Policy Gradients

DGX agent

arXiv:2607.05272v1 Announce Type: cross Abstract: Inference serving systems must balance throughput and latency under bursty, heterogeneous workloads, yet the industry standard remains static batching

safetyarxiv-cs-ai
7 Jul 2026
Local Ai

AdaptiveSD A Stability-Aware, Runtime-Adaptive Speculative Decoding Framework with Multi-Policy Orchestration for CPU-Constrained LLM Inference

DGX agent

arXiv:2607.03876v1 Announce Type: new Abstract: With the rise of small quantized GGUF-based language models and their increasing use for on-device inference tasks, we have seen the growing need for an

local-aiarxiv-cs-lg
7 Jul 2026
Model Releases

EM3M: An Electron Micrograph Dataset for Microstructural Segmentation and Generation

DGX agent

arXiv:2508.16239v2 Announce Type: replace Abstract: Quantitative microstructural characterization is fundamental to materials science, and electron micrographs (EMs) provide indispensable high-resolut

model-releasesarxiv-cs-cv
7 Jul 2026
Applications

From Geometric Labels to Semantic Understanding of Indoor Building Components Using Multimodal Large Language Models

DGX agent

arXiv:2607.03661v1 Announce Type: new Abstract: Point cloud-based understanding has become an important enabler for facility operation and maintenance involving indoor building components. However, ex

applicationsarxiv-cs-cv
7 Jul 2026
Research

Inverse Design of Metainterfaces for Static Friction Control: Beyond the Hertzian Limit

DGX agent

arXiv:2605.11012v2 Announce Type: cross Abstract: Programming the static friction of mechanical interfaces is critical for soft robotics, haptics, and precision gripping. Static friction is governed b

researcharxiv-cs-lg
7 Jul 2026
Model Releases

Metronome: Bound the Cache, Keep the Beat for Real-Time Interaction Model Serving

DGX agent

arXiv:2607.02640v1 Announce Type: cross Abstract: Real-time interaction models -- Moshi, MiniCPM-o, Qwen-Omni -- turn serving into a periodic real-time task: on every frame a session ingests streaming

model-releasesarxiv-cs-ai
7 Jul 2026
Hardware

MLSYSIM: First-Principles Infrastructure Modeling for Machine Learning Systems

DGX agent

arXiv:2607.02558v1 Announce Type: cross Abstract: As machine learning shifts from laboratory curiosity to critical infrastructure, the systems that sustain it span an extraordinary range, from sub-mil

hardwarearxiv-cs-lg
7 Jul 2026
Agents

SMOCS: A Streaming Framework for Simplified Deployment, Monitoring, and Optimization of ML Systems in Production

DGX agent

arXiv:2607.02731v1 Announce Type: cross Abstract: Machine learning has demonstrated significant potential for real-time monitoring, optimization, and control of scientific facilities. However, deployi

agentsarxiv-cs-ai
7 Jul 2026
Model Releases

Streaming Model Cascades for Semantic SQL

DGX agent

arXiv:2604.00660v2 Announce Type: replace-cross Abstract: Modern data warehouses extend SQL with semantic operators that invoke large language models on each qualifying row, making per-row inference o

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

Safe and Adaptive Cloud Healing: Verifying LLM-Generated Recovery Plans with a Neural-Symbolic World Model

DGX agent

arXiv:2607.01595v1 Announce Type: new Abstract: As the scale and complexity of cloud-based AI systems continue to escalate, ensuring service reliability through rapid fault detection and adaptive reco

safetyarxiv-cs-ai
3 Jul 2026
Research

ThreadWeaver: Adaptive Threading for Efficient Parallel Reasoning in Language Models

DGX agent

arXiv:2512.07843v2 Announce Type: replace-cross Abstract: Scaling inference-time computation has enabled Large Language Models (LLMs) to achieve strong reasoning performance, but their inherently sequ

researcharxiv-cs-ai
3 Jul 2026
Agents

A Single Rewrite Suffices: Empirical Lessons from Production Skill Description Optimization

DGX agent

arXiv:2606.30775v1 Announce Type: cross Abstract: Enterprise AI agents route user queries to specialized skills by matching queries against natural language skill descriptions. When two skills share o

agentsarxiv-cs-ai
1 Jul 2026
Safety

QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents

DGX agent

arXiv:2606.32034v1 Announce Type: cross Abstract: LLM agents increasingly act over long horizons, where a single trajectory can contain hundreds or thousands of actions. In these settings, outcome-onl

safetyarxiv-cs-ai
1 Jul 2026
Agents

An AI Security Agent for Banking: Multi-Vector Fraud and AML Detection Across Retail and Corporate Accounts

DGX agent

arXiv:2606.17555v2 Announce Type: replace-cross Abstract: Banks face two threat families with fundamentally different detection requirements: signature-based fraud (card-not-present attacks, account t

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

Defeat Devices in AI Systems

DGX agent

arXiv:2606.28863v1 Announce Type: cross Abstract: AI systems increasingly exhibit behavior that differs systematically between evaluation and deployment contexts. Alignment faking, sandbagging, benchm

model-releasesarxiv-cs-ai
30 Jun 2026
Hardware

HARD-KV: Head-Adaptive Regularization for Decoding-time KV Compression

DGX agent

arXiv:2606.28831v1 Announce Type: cross Abstract: Long-context LLM inference faces a fundamental conflict: head-adaptive compression algorithms (e.g., Top-p nucleus sampling) offer superior accuracy b

hardwarearxiv-cs-ai
30 Jun 2026
Research

Harvesting AI Computation at the Edge via Generic Approximation

DGX agent

arXiv:2606.29518v1 Announce Type: cross Abstract: With the widespread adoption of AI in various IoT scenarios such as smart sensing and processing, AI chips have become a common component at the edge.

researcharxiv-cs-lg
30 Jun 2026
Research

Structural Certification for Reliable Physical Design with Language Models

DGX agent

arXiv:2606.30107v1 Announce Type: new Abstract: An unreliable language model can be made to produce reliable physical designs if the authority to assert is moved out of the model: the model proposes,

researcharxiv-cs-ai
30 Jun 2026
Safety

The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning

DGX agent

arXiv:2606.29526v1 Announce Type: new Abstract: Reinforcement learning (RL) has gained growing attention in large language model (LLM) post-training, yet RL training remains fragile and can suffer fro

safetyarxiv-cs-lg
30 Jun 2026
Safety

The Undecidability of Artificial General Intelligence (AGI) Alignment

DGX agent

arXiv:2606.28639v1 Announce Type: cross Abstract: This article establishes the foundational mathematical limits of Artificial General Intelligence (AGI) safety, proving that the core barrier is not th

safetyarxiv-cs-ai
30 Jun 2026
Safety

MetaBreak: Jailbreaking Online LLM Services via Special Token Manipulation

DGX agent

arXiv:2510.10271v2 Announce Type: replace-cross Abstract: Unlike regular tokens derived from existing text corpora, special tokens are artificially created to annotate structured conversations during

safetyarxiv-cs-ai
29 Jun 2026
Model Releases

Boundary-Aware Context Grounding for A Low-Channel EEG Agent

DGX agent

arXiv:2606.26519v1 Announce Type: new Abstract: Large language models (LLMs) can make scientific software easier to use. However, a general model does not automatically know which measurements a parti

model-releasesarxiv-cs-ai
26 Jun 2026
Applications

Investigating LLM's Problem Solving Capability -- a Study on Statics Questions

DGX agent

arXiv:2606.26103v1 Announce Type: cross Abstract: Large Language Models (LLMs) have rapidly influenced many aspects of society, particularly education, due to their demonstrated ability to complete as

applicationsarxiv-cs-ai
26 Jun 2026
Model Releases

NuclearQAv2: A Structured Benchmark for Evaluating Domain-Science Competence in Large Language Models

DGX agent

arXiv:2606.27047v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong performance across a wide range of tasks, but ensuring their reliability in highly technical dom

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Scalable AI-assisted Workflow Management for Detector Design Optimization Using Distributed Computing

DGX agent

arXiv:2603.30014v2 Announce Type: replace-cross Abstract: The Production and Distributed Analysis (PanDA) system, originally developed for the ATLAS experiment at the CERN Large Hadron Collider (LHC),

model-releasesarxiv-cs-ai
26 Jun 2026
Safety

MAPL: Multi-Objective Preference Learning for Robot Locomotion

DGX agent

arXiv:2606.25398v1 Announce Type: new Abstract: Reward design remains a major bottleneck in reinforcement learning for robot locomotion, where successful policies often depend on carefully tuned, task

safetyarxiv-cs-ro
25 Jun 2026
Model Releases

Age of LLM: A Strategic 1v1 Benchmark for Reasoning, Diplomacy and Reliability of Large Language Models under Fog of War

DGX agent

arXiv:2606.24391v1 Announce Type: new Abstract: We introduce Age of LLM, a turn-based 1v1 benchmark in which two LLMs face off on a 13x7 grid to destroy the enemy base. Three stressors are deliberate:

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

AI-PAVE-Br: Leveraging Large Language Models for Enhanced Product Attribute Value Extraction through a Golden Set Approach

DGX agent

arXiv:2606.24655v1 Announce Type: cross Abstract: The explosive growth and complexity of product data within the dynamic Brazilian e-commerce landscape demand robust and specialized methods for struct

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

BIM-Edit: Benchmarking Large Language Models for IFC-Based Building Information Modeling

DGX agent

arXiv:2606.20146v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied to computer-aided design (CAD) to generate design artifacts from textual instructions. In engi

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

DeepBD: A Grounded Agentic Workflow for Variant Prioritization and Diagnosis of Genetic Birth Defects

DGX agent

arXiv:2606.24779v1 Announce Type: cross Abstract: Birth defects are a major cause of fetal loss, neonatal morbidity and long-term disability. In the subset with suspected genetic etiologies, exome and

model-releasesarxiv-cs-ai
24 Jun 2026
Safety

DeformX: A Versatile Co-Simulation Framework for Deformable Linear Objects

DGX agent

arXiv:2606.22116v1 Announce Type: new Abstract: Deformable linear objects (DLOs) such as wires, cables, and ropes are common in robotic manipulation tasks, yet simulating them with both visual realism

safetyarxiv-cs-ro
23 Jun 2026
Model Releases

A Five-Plane Reference Architecture for Runtime Governance of Production AI Agents

DGX agent

arXiv:2606.12320v1 Announce Type: new Abstract: Enterprise security was built to govern data boundaries: the protected surface was data at rest and in transit, and the controls -- access control, data

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Human-Guided Agentic AI for Multimodal Clinical Prediction: Lessons from the AgentDS Healthcare Benchmark

DGX agent

arXiv:2602.19502v2 Announce Type: replace Abstract: Agentic AI systems are increasingly capable of autonomous data science workflows, yet clinical prediction tasks demand domain expertise that purely

model-releasesarxiv-cs-ai
11 Jun 2026
Applications

Trace2Policy: From Expert Behavior Traces to Self-Evolving Decision Agents

DGX agent

arXiv:2606.10457v1 Announce Type: new Abstract: Decision rules that enterprise experts apply tacitly -- in auditing, compliance, and contract review -- can be systematically recovered and improved thr

applicationsarxiv-cs-ai
10 Jun 2026
Model Releases

Real-Time Industrial Defect Detection on Edge Hardware Using Fine-Tuned YOLOv8: A Systematic Benchmark on the NEU Surface Defect Database and MVTec AD with Automotive & Battery Manufacturing Extensions

DGX agent

arXiv:2606.07659v1 Announce Type: new Abstract: Automated surface defect detection is critical for ensuring rigorous quality control in high-speed manufacturing environments. While deep learning model

model-releasesarxiv-cs-cv
9 Jun 2026
Agents

Structuring agentic AI for HPC code modernization

DGX agent

arXiv:2606.08710v1 Announce Type: cross Abstract: Modernization of legacy scientific codes is often necessary to keep up with the ever-evolving changes in the compute resource ecosystem. Parallelizati

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

vla.cpp: A Unified Inference Runtime for Vision-Language-Action Models

DGX agent

arXiv:2606.08094v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies are typically shipped as Python/PyTorch stacks that assume a workstation-class GPU, a mismatch for the hardware

model-releasesarxiv-cs-ai
9 Jun 2026
Applications

Design Once, Deploy at Scale: Template-Driven ML Development for Large Model Ecosystems

DGX agent

arXiv:2603.24963v3 Announce Type: replace Abstract: Modern computational advertising platforms typically rely on recommendation systems to predict user responses, such as click-through rates, conversi

applicationsarxiv-cs-ai
8 Jun 2026
Research

Multimodal Sexism Identification and Characterization using Large Language Models and Gradient Boosting

DGX agent

arXiv:2606.05997v1 Announce Type: new Abstract: We present the AILS-NTUA submission to the EXIST 2026 Lab at CLEF, addressing multimodal sexism identification and characterization in memes (Task 2) an

researcharxiv-cs-cv
5 Jun 2026
Model Releases

Evaluating Zero-Shot and One-Shot Adaptation of Small Language Models in Leader-Follower Interaction

DGX agent

arXiv:2602.23312v3 Announce Type: replace-cross Abstract: Leader-follower interaction is an important paradigm in human-robot interaction (HRI). Yet, assigning roles in real time remains challenging f

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

From Prompt to Process: a Process Taxonomy and Comparative Assessment of Frameworks Supporting AI Software Development Agents

DGX agent

arXiv:2606.04967v1 Announce Type: cross Abstract: AI tools for programming are no longer just autocomplete or chat assistants: they organize themselves as development frameworks, with process, roles,

agentsarxiv-cs-ai
4 Jun 2026
Agents

EvoTrainer: Co-Evolving LLM Policies and Training Harnesses for Autonomous Agentic Reinforcement Learning

DGX agent

arXiv:2606.03108v1 Announce Type: new Abstract: Autonomous LLM training is often framed as recipe search, which leaves the training harness largely static. This limitation sharpens in agentic RL, wher

agentsarxiv-cs-ai
3 Jun 2026
← Previous
1…2122232425…66
Next →