AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “engineering”

GridTimelineEvolution
3,145 results
Research

Test-Time Optimization of Physical Query Plans with LLMs

DGX agent

arXiv:2602.10387v2 Announce Type: replace-cross Abstract: Traditional query optimization relies on cost-based optimizers that estimate execution cost (e.g., runtime, memory, and I/O) using predefined

researcharxiv-cs-ai
3 Jun 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

From Demonstrations to Rewards: Test-Time Prompt Optimization for VLM Reward Models

DGX agent

arXiv:2606.00083v1 Announce Type: cross Abstract: Reinforcement learning relies on accurate reward functions, which are often hand-crafted or even unavailable in real-world applications, such as robot

safetyarxiv-cs-ai
2 Jun 2026
Research

HyperVQ: Enabling Hyperprior Entropy Modeling for VQ-Based Generative Image Compression

DGX agent

arXiv:2512.07192v2 Announce Type: replace Abstract: Vector Quantization (VQ) based generative image compression has achieved remarkable perceptual quality. However, existing VQ codecs suffer from two

researcharxiv-cs-cv
2 Jun 2026
Agents

Learning to Construct Practical Agentic Systems

DGX agent

arXiv:2606.00189v1 Announce Type: cross Abstract: Automated design and optimization of agentic LLM-based systems leads to sophisticated systems that substantially improve result quality over off-the-s

agentsarxiv-cs-ai
2 Jun 2026
Agents

PairedGTA: Generating Driving Datasets for Controlled Photometric Shift Analysis

DGX agent

arXiv:2606.01192v1 Announce Type: new Abstract: Evaluating the performance of visual perception systems for autonomous driving is essential to ensure reliable operation across diverse environmental sc

agentsarxiv-cs-cv
2 Jun 2026
Safety

Post-Deterministic Distributed Systems: A New Foundation for Trustworthy Autonomous Infrastructure

DGX agent

arXiv:2606.01722v1 Announce Type: cross Abstract: For decades, distributed systems have typically assumed that correct participants execute protocol-specified behavior with stable, externally defined,

safetyarxiv-cs-ai
2 Jun 2026
Local Ai

scicode-lint: Detecting Methodology Bugs in Scientific Python Code with LLM-Generated Patterns

DGX agent

arXiv:2603.17893v2 Announce Type: replace-cross Abstract: Methodology bugs in scientific Python code produce plausible but incorrect results that traditional linters and static analysis tools cannot d

local-aiarxiv-cs-ai
2 Jun 2026
Agents

Towards Interactive Video World Modeling: Frontiers, Challenges, Benchmarks, and Future Trends

DGX agent

arXiv:2606.01164v1 Announce Type: new Abstract: With rapid development of large language models and diffusion-based content generation, world modeling has attracted increasing research attention, bene

agentsarxiv-cs-cv
2 Jun 2026
Hardware

Deterministic Inference across Tensor Parallel Sizes That Eliminates Training-Inference Mismatch

DGX agent

arXiv:2511.17826v2 Announce Type: replace-cross Abstract: Deterministic inference is increasingly critical for large language model (LLM) applications such as LLM-as-a-judge evaluation, multi-agent sy

hardwarearxiv-cs-cl
1 Jun 2026
Safety

BEAMS: Benchmarking and Evaluating AI for Modeling and Simulation

DGX agent

arXiv:2605.28994v1 Announce Type: new Abstract: AI tools to support real world decision making must be able to build simulation models that inform their recommendations and render them interpretable.

safetyarxiv-cs-ai
29 May 2026
Model Releases

LLM-Evolved Domain-Independent Heuristics for Symbolic AI Planning

DGX agent

arXiv:2605.29649v1 Announce Type: new Abstract: Heuristic search is the dominant paradigm in symbolic AI planning, and the strongest heuristics are the result of decades of work by planning researcher

model-releasesarxiv-cs-ai
29 May 2026
Research

Singularity-free dynamical invariants-based quantum control

DGX agent

arXiv:2510.15340v2 Announce Type: replace-cross Abstract: State preparation is a cornerstone of quantum technologies, underpinning applications in computation, communication, and sensing. Its importan

researcharxiv-cs-lg
29 May 2026
Model Releases

Temporal Stability and Few-Shot Prompting in Math Task Assessment

DGX agent

arXiv:2605.30151v1 Announce Type: new Abstract: As AI tools become increasingly integrated into educational contexts, questions arise about both their stability over time and their responsiveness to p

model-releasesarxiv-cs-ai
29 May 2026
Agents

AIBuildAI-2: A Knowledge-Enhanced Agent for Automatically Building AI Models

DGX agent

arXiv:2605.27873v1 Announce Type: new Abstract: AI models underpin data-centric applications from image and text processing to scientific discovery in biology, physics, and chemistry. Yet developing t

agentsarxiv-cs-ai
28 May 2026
Applications

OphIn-500K: Curating Web-Scale Visual Instructions for Scaling Ophthalmic Multimodal Large Language Models

DGX agent

arXiv:2605.27916v1 Announce Type: cross Abstract: The advancement of general medical Multimodal Large Language Models (MLLMs) has shown great potential for building conversational assistants to suppor

applicationsarxiv-cs-cl
28 May 2026
Agents

QUACK: Questioning, Understanding, and Auditing Communicated Knowledge in Multimodal Social Deduction Agents

DGX agent

arXiv:2605.27068v1 Announce Type: cross Abstract: Social deduction games have become a popular testbed for probing reasoning, deception, coordination, and belief modeling in Large Language Model (LLM)

agentsarxiv-cs-ai
27 May 2026
Tutorials

An Empirical Evaluation of LLM-Generated Code Security Across Prompting Methods

DGX agent

arXiv:2605.24298v1 Announce Type: cross Abstract: The growing use of Large Language Models (LLMs) for automated code generation has enhanced software development efficiency, but often at the cost of s

tutorialsarxiv-cs-ai
26 May 2026
Safety

Cultivating Machine Intelligence: The OMEGA Shift from Top-Down Optimization to Autopoietic Cognitive Ecologies

DGX agent

arXiv:2605.25062v1 Announce Type: cross Abstract: The dominant artificial intelligence paradigm trains neural architectures via gradient descent against proxy objectives and reinforcement learning fro

safetyarxiv-cs-ai
26 May 2026
Agents

Decoding ML Decision: An Agentic Reasoning Framework for Large-Scale Ranking System

DGX agent

arXiv:2602.18640v2 Announce Type: replace Abstract: Modern large-scale ranking systems operate within a sophisticated landscape of competing objectives, operational constraints, and evolving product r

agentsarxiv-cs-ai
26 May 2026
Agents

PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback

DGX agent

arXiv:2605.24775v1 Announce Type: new Abstract: Operating LLMs as coordinated multi-agent research systems over multi-hour runs surfaces failure modes that single-shot evaluation cannot: upstream prov

agentsarxiv-cs-ai
26 May 2026
Agents

AI Assurance: A Comprehensive Testing Strategy for Enterprise AI Systems

DGX agent

arXiv:2605.23459v1 Announce Type: cross Abstract: Enterprise AI systems, built on large language models, retrieval pipelines and autonomous agents, introduce a class of risks that traditional software

agentsarxiv-cs-ai
25 May 2026
Model Releases

One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents

DGX agent

arXiv:2605.23652v1 Announce Type: new Abstract: On a 300-persona life-simulation benchmark, pcsp achieves compositional zero-shot persona identification up to 17x above chance, Spearman rho approx 0.7

model-releasesarxiv-cs-ai
25 May 2026
Safety

Mapping Tomato Cropping Systems in California Using AlphaEarth Geospatial Embeddings and Deep Learning Analysis

DGX agent

arXiv:2605.21804v1 Announce Type: cross Abstract: Field-scale crop maps support supply-chain forecasting and policy, yet statewide crop identification still often depends on retrospective surveys or r

safetyarxiv-cs-cv
22 May 2026
Research

A Mechanistic Study of Tabular Foundation Models

DGX agent

arXiv:2605.21288v1 Announce Type: new Abstract: Tabular foundation models with different architectures converge in accuracy across a range of classification and regression tasks. This raises questions

researcharxiv-cs-lg
21 May 2026
Model Releases

Leveraging Vision-Language Models to Detect Attention in Educational Videos

DGX agent

arXiv:2605.20211v1 Announce Type: new Abstract: Educational videos are a cornerstone of remote and blended learning. However, learners' fluctuating attention remains a significant barrier to effective

model-releasesarxiv-cs-cv
21 May 2026
Agents

Retrieval-Augmented Code Generation: A Survey with Focus on Repository-Level Approaches

DGX agent

arXiv:2510.04905v3 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have significantly improved automated code generation. While existing approaches have achieved

agentsarxiv-cs-cl
21 May 2026
Safety

SUGAR: A Scalable Human-Video-Driven Generalizable Humanoid Loco-Manipulation Learning Framework

DGX agent

arXiv:2605.20373v1 Announce Type: cross Abstract: Building humanoid robots capable of generalizable whole-body loco-manipulation in the real world remains a fundamental challenge. Existing methods eit

safetyarxiv-cs-cv
21 May 2026
Model Releases

Understanding and Improving Communication Performance in Multi-node LLM Inference

DGX agent

arXiv:2511.09557v4 Announce Type: replace-cross Abstract: As large language models (LLMs) continue to grow in size, distributed inference has become increasingly important. Model-parallel strategies m

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Entry-level guide to the use of large language models for medical research

DGX agent

arXiv:2410.18856v4 Announce Type: replace Abstract: Frontier large language models (LLMs), such as GPT-5, Claude 4.5, Gemini 3, Llama 4, and DeepSeek-R1, represent a transformative class of AI tools c

model-releasesarxiv-cs-ai
20 May 2026
Agents

Operationalising Artificial Intelligence Bills of Materials (AIBOMs) for Verifiable AI Provenance and Lifecycle Assurance

DGX agent

arXiv:2605.19755v1 Announce Type: cross Abstract: Artificial Intelligence (AI) systems are increasingly dependent on complex, multi-layered software supply chains that introduce challenges for reprodu

agentsarxiv-cs-ai
20 May 2026
Model Releases

BlendedNet++: A dataset and benchmark for field-resolved aerodynamics and inverse design of blended wing body aircraft

DGX agent

arXiv:2512.03280v2 Announce Type: replace-cross Abstract: The conceptual design of Blended Wing Body (BWB) aircraft is often constrained by the high computational cost of resolving complex aerodynamic

model-releasesarxiv-cs-ai
19 May 2026
Safety

Code as Agent Harness

DGX agent

arXiv:2605.18747v1 Announce Type: cross Abstract: Recent large language models (LLMs) have demonstrated strong capabilities in understanding and generating code, from competitive programming to reposi

safetyarxiv-cs-ai
19 May 2026
Research

GeoSym127K: Scalable Symbolically-verifiable Synthesis for Multimodal Geometric Reasoning

DGX agent

arXiv:2605.16371v1 Announce Type: cross Abstract: Large Multimodal Models (LMMs) often struggle with geometric reasoning due to visual hallucinations and a lack of mathematically precise Chain-of-Thou

researcharxiv-cs-ai
19 May 2026
Research

Learning to Look Benign: Targeted Evasion of Malware Detectors via API Import Injection

DGX agent

arXiv:2605.18624v1 Announce Type: cross Abstract: Machine learning-based malware detectors are widely deployed in antivirus and endpoint detection systems, yet their reliance on static features makes

researcharxiv-cs-lg
19 May 2026
Research

Response-free item difficulty modelling for multiple-choice items with fine-tuned transformers: Component-wise representation and multi-task learning

DGX agent

arXiv:2605.16991v1 Announce Type: cross Abstract: Response-free item difficulty modelling promises to reduce reliance on response-based calibration but is intrinsically difficult on reading-comprehens

researcharxiv-cs-ai
19 May 2026
Hardware

Systematic Optimization of Real-Time Diffusion Model Inference on Apple M3 Ultra

DGX agent

arXiv:2605.16259v1 Announce Type: cross Abstract: While real-time image generation using diffusion models has advanced rapidly on NVIDIA GPUs, systematic optimization research on non-CUDA platforms su

hardwarearxiv-cs-ai
19 May 2026
Hardware

A Few GPUs, A Whole Lotta Scale: Faithful LLM Training Emulation with PrismLLM

DGX agent

arXiv:2605.15617v1 Announce Type: cross Abstract: Large language model (LLM) training today runs on clusters spanning thousands of GPUs. While this scale enables rapid model advances, developing, debu

hardwarearxiv-cs-ai
18 May 2026
Safety

CTF4Nuclear: Common Task Framework for Nuclear Fission and Fusion Models

DGX agent

arXiv:2605.15549v1 Announce Type: cross Abstract: The demand for clean energy is ever increasing, with new nuclear technologies presenting a complementary solution to renewable energies. However, desi

safetyarxiv-cs-ai
18 May 2026
Applications

Layer-wise Derivative Controlled Networks

DGX agent

arXiv:2605.15463v1 Announce Type: new Abstract: As machine learning models grow in complexity, they increasingly struggle with three conflicting demands: the need for high accuracy, the requirement fo

applicationsarxiv-cs-lg
18 May 2026
Agents

PRISM: Prompt Reliability via Iterative Simulation and Monitoring for Enterprise Conversational AI

DGX agent

arXiv:2605.15665v1 Announce Type: new Abstract: Deploying large language model (LLM)-driven conversational agents in enterprise settings requires prompts that are simultaneously correct at launch and

agentsarxiv-cs-ai
18 May 2026
Research

Context-Aware Web Attack Detection in Open-Source SIEM Systems via MITRE ATT&CK-Enriched Behavioral Profiling

DGX agent

arXiv:2605.13337v1 Announce Type: cross Abstract: Security Information and Event Management (SIEM) systems aggregate log data from heterogeneous sources to detect coordinated attacks. Traditional rule

researcharxiv-cs-lg
14 May 2026
Safety

MinT: Managed Infrastructure for Training and Serving Millions of LLMs

DGX agent

arXiv:2605.13779v1 Announce Type: cross Abstract: We present MindLab Toolkit (MinT), a managed infrastructure system for Low-Rank Adaptation (LoRA) post-training and online serving. MinT targets a set

safetyarxiv-cs-ai
14 May 2026
Research

Self-CriTeach: LLM Self-Teaching and Self-Critiquing for Improving Robotic Planning via Automated Domain Generation

DGX agent

arXiv:2509.21543v3 Announce Type: replace Abstract: Large Language Models (LLMs) have shown strong promise for robotic task planning, particularly through the automatic generation of symbolic planning

researcharxiv-cs-ro
14 May 2026
Model Releases

Crash Assessment via Mesh-Based Graph Neural Networks and Physics-Aware Attention

DGX agent

arXiv:2605.11784v1 Announce Type: cross Abstract: Full-vehicle crash simulations are computationally expensive, limiting their use in iterative design exploration. This work investigates learned hybri

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Nautilus: From One Prompt to Plug-and-Play Robot Learning

DGX agent

arXiv:2605.11665v1 Announce Type: new Abstract: Robot learning research is fragmented across policy families, benchmark suites, and real robots; each implementation is entangled with the others in a c

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

AgentCollabBench: Diagnosing When Good Agents Make Bad Collaborators

DGX agent

arXiv:2605.08647v1 Announce Type: cross Abstract: Multi-agent systems achieve state-of-the-art outcomes through peer collaboration. However, when an agent in the pipeline silently drops a constraint,

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

BenchCAD: A Comprehensive, Industry-Standard Benchmark for Programmatic CAD

DGX agent

arXiv:2605.10865v1 Announce Type: new Abstract: Industrial Computer-Aided Design (CAD) code generation requires models to produce executable parametric programs from visual or textual inputs. Beyond r

model-releasesarxiv-cs-ai
12 May 2026
Safety

Big AI is accelerating the metacrisis: What can we do?

DGX agent

arXiv:2512.24863v2 Announce Type: replace-cross Abstract: The world is in the grip of ecological, meaning, and language crises that are converging into a metacrisis. Big AI is accelerating them all. L

safetyarxiv-cs-ai
12 May 2026
← Previous
1…2223242526…66
Next →