AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,134 results
Model Releases

Same Question, Different Answer? Measuring and Mitigating Prompt Privilege for Equitable AI Access

DGX agent

arXiv:2608.08942v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into healthcare, education, public services, and everyday decision making. They should provide

model-releasesarxiv-cs-cl
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SkillConsist: Detecting Inconsistencies in Agent Skills via Bidirectional Graph Alignment

DGX agent

arXiv:2608.07639v1 Announce Type: cross Abstract: Agent Skills provide reusable capabilities to LLM agents. Agent Skill inconsistencies can expose undisclosed dangerous behavior or cause wrong Skill s

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Stealing Reasoning Traces from Proprietary LLM APIs

DGX agent

arXiv:2608.09867v1 Announce Type: cross Abstract: Leading large language model providers now conceal their models' step-by-step reasoning, or chain-of-thought, to protect intellectual property and lim

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

SuperCoder: Assembly Program Superoptimization with Large Language Models

DGX agent

arXiv:2505.11480v4 Announce Type: replace-cross Abstract: Superoptimization is the task of transforming a program into a faster one, and ideally the very fastest possible one, while preserving its inp

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SurveyReview: A Reviewer-Aligned Benchmark for Survey Evaluators

DGX agent

arXiv:2608.07641v1 Announce Type: new Abstract: The rapid advancement of large language models has transformed survey writing from a months-long manual effort into an automated process. As generation

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

The Evolution of Mixture-of-Experts Architectures in Large Language Models: Routing, Topology, Load Balancing, and Expert Parallelism

DGX agent

arXiv:2608.08650v1 Announce Type: new Abstract: Mixture-of-Experts models increase parameter capacity while keeping the computation activated by each token bounded, but their architectural evolution c

model-releasesarxiv-cs-cl
11 Aug 2026
Safety

The Voiceprint Fallacy: Why Voices Are Not Unique Biometric Imprints

DGX agent

arXiv:2608.07980v1 Announce Type: cross Abstract: In recent years, the term voiceprint has regained attention, particularly in technological applications and policy-making contexts, often carrying the

safetyarxiv-cs-cl
11 Aug 2026
Applications

Three Generations of Healthcare IT: From the Digital Record to the Computable Care Process

DGX agent

arXiv:2608.08806v1 Announce Type: new Abstract: Objective. Healthcare IT is usually organized by the technologies it adopts. We instead organize it by the unit of information a system makes computable

applicationsarxiv-cs-ai
11 Aug 2026
Safety

ToolUniverse: An open platform for democratizing AI scientists

DGX agent

arXiv:2509.23426v3 Announce Type: replace Abstract: AI scientists are emerging computational systems that serve as collaborative partners in discovery. These systems remain difficult to build because

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Towards an LLM-based method for quantifying the sexual content in song lyrics

DGX agent

arXiv:2608.08885v1 Announce Type: cross Abstract: Reggaeton is one of the most widely consumed music genres in the world, and its lyrics are commonly regarded as highly sexualized. This claim rests mo

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Towards Expert-level Medical AI for Real-time Video Consultations

DGX agent

arXiv:2608.09861v1 Announce Type: new Abstract: Audio-visual interaction is the standard for patient-physician consultations, enabling natural communication and effective assessment of illness through

model-releasesarxiv-cs-ai
11 Aug 2026
Hardware

verdi: retrieval is not transfer for continual world model optimization

DGX agent

arXiv:2608.09537v1 Announce Type: new Abstract: Foundation world models have made remarkable progress in planning, simulation, and embodied intelligence. However, optimizing a pretrained world model t

hardwarearxiv-cs-ai
11 Aug 2026
Model Releases

Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation

DGX agent

arXiv:2608.07762v1 Announce Type: new Abstract: LLM benchmarks can build an organization's reputation and attract customers, but only when results are transparent and verifiable. Unverified claims tha

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

A Multi-Agent Framework for Automated Coarse-Grained Molecular Dynamics of Polymers

DGX agent

arXiv:2608.06694v1 Announce Type: new Abstract: Coarse-grained (CG) molecular dynamics extends polymer simulation beyond the scales accessible to all-atom (AA) methods, but bottom-up CG modeling is la

model-releasesarxiv-cs-ai
10 Aug 2026
Agents

ADIAS: Automated Design of Interactive Agentic Systems

DGX agent

arXiv:2608.06410v1 Announce Type: new Abstract: Automated agent design improves agent harnesses through iterative revision, evaluation, and feedback summarization. Existing methods are largely candida

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

CADSpotting: Robust Panoptic Symbol Spotting on Large-Scale CAD Drawings

DGX agent

arXiv:2412.07377v5 Announce Type: replace Abstract: We introduce CADSpotting, an effective method for panoptic symbol spotting in large-scale architectural CAD drawings. Existing approaches often stru

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Critical Acclaim Orientation in Large Language Models: Evidence from Film Preference Elicitation

DGX agent

arXiv:2608.06955v1 Announce Type: new Abstract: Large language models (LLMs) are trained on corpora that contain expressions of human judgment about films, books, music, and more. Yet whether LLMs sys

model-releasesarxiv-cs-ai
10 Aug 2026
Local Ai

Dual-Node NVIDIA DGX Spark over Tailscale: A Remote-Access Testbed for Distributed LLM Training and Cyber-Threat-Intelligence Fine-Tuning

DGX agent

arXiv:2608.07226v1 Announce Type: cross Abstract: Compact AI systems make local language-model experimentation increasingly accessible, yet practical evidence for multi-node training on desktop-class

local-aiarxiv-cs-lg
10 Aug 2026
Model Releases

ED-CSP: Crystal Structure Prediction from Electron Diffraction

DGX agent

arXiv:2608.06448v1 Announce Type: cross Abstract: Recovering a periodic 3D crystal structure from sparse, unindexed electron diffraction (ED) observations is a challenging generative inverse problem.

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Evaluating XAI Support From A Hierarchical Reinforcement Learning Policy in Human-Agent Collaboration

DGX agent

arXiv:2608.06381v1 Announce Type: cross Abstract: Explainable AI (XAI) has shown promise for human-agent collaboration, yet results rely on hand-crafted policies in custom environments, limiting gener

model-releasesarxiv-cs-ai
10 Aug 2026
Agents

Kimi K2.5: Visual Agentic Intelligence

DGX agent

arXiv:2602.02276v2 Announce Type: replace-cross Abstract: We introduce Kimi K2.5, an open-source multimodal agentic model designed to advance general agentic intelligence. K2.5 emphasizes the joint op

agentsarxiv-cs-ai
10 Aug 2026
Safety

LMM Modality Transfer: A Pre-requisite for Autonomous GIS Agents

DGX agent

arXiv:2608.06948v1 Announce Type: new Abstract: AI models are becoming increasingly adept at understanding and processing spatial information, thereby facilitating agentic problem-solving in spatial t

safetyarxiv-cs-ai
10 Aug 2026
Safety

MaskFlow: Precise, Consistent and Seamless Regional Image Editing

DGX agent

arXiv:2608.06929v1 Announce Type: cross Abstract: Regional image editing has attracted considerable attention for its spatial controllability. Although instruction-based and mask-reference-based editi

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

MirrorWorld: Taming Video Diffusion Models for Mirror Reflection Generation

DGX agent

arXiv:2608.07463v1 Announce Type: new Abstract: Recent advances in video diffusion models (VDMs) have enabled high-fidelity video synthesis. However, generating mirror reflections remains challenging

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Panoramic Multimodal Semantic Occupancy Prediction for Quadruped Robots

DGX agent

arXiv:2603.13108v2 Announce Type: replace-cross Abstract: Panoramic imagery provides holistic 360{eg} visual coverage for environmental perception in quadruped robots. However, existing occupancy pred

model-releasesarxiv-cs-cv
10 Aug 2026
Applications

Sharding Prevents LLM Oversight Failures and Adversarial Exploitation

DGX agent

arXiv:2608.06422v1 Announce Type: new Abstract: Giving an LLM judge more compute does not necessarily make it check more requirements. When one call must return many verdicts, some decisions become we

applicationsarxiv-cs-lg
10 Aug 2026
Agents

SyncSBC: Decentralized Swarm Behavior Prediction for Synchronized Autonomous Control

DGX agent

arXiv:2608.06587v1 Announce Type: cross Abstract: Robot swarms utilize many independent limited-sensing agents to produce complex emergent behaviors without requiring centralized control. However, lit

agentsarxiv-cs-ai
10 Aug 2026
Safety

TA-RAG: Tone Awareness as a Design Imperative for Retrieval-Augmented Generation

DGX agent

arXiv:2608.06672v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become a robust architecture for grounding large language models (LLMs) in trusted knowledge. However, standard

safetyarxiv-cs-cl
10 Aug 2026
Agents

Towards Assurance Closure in AI-Native Large-Scale Agile Software Development

DGX agent

arXiv:2608.07317v1 Announce Type: cross Abstract: The AI-Native Manifesto envisions large-scale agile software development in which humans increasingly govern intent, risk, and exceptions while agents

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

An Axiomatic Benchmark for Evaluation of Scientific Novelty Metrics

DGX agent

arXiv:2604.15145v2 Announce Type: replace Abstract: The rigorous evaluation of the novelty of a scientific paper is, even for human scientists, a challenging task. With the increasing interest in AI s

model-releasesarxiv-cs-ai
7 Aug 2026
Agents

ASTELD: A Six-Axis Classification Framework for Autonomous AI Agents - Design, Evaluation, and an OpenClaw Case Study

DGX agent

arXiv:2608.05201v1 Announce Type: cross Abstract: Autonomous AI agent platforms differ substantially in architecture, security, tool integration, execution, autonomy, and deployment, yet the field lac

agentsarxiv-cs-ai
7 Aug 2026
Safety

Clinical Communication Processing with Models Trained on LLM-Generated Synthetic Data: A Structured Survey and Novel Application Case Studies

DGX agent

arXiv:2608.05993v1 Announce Type: new Abstract: Much clinical value is conveyed not through structured records but through communication: exchanges in which patients describe symptoms, clinicians reas

safetyarxiv-cs-cl
7 Aug 2026
Model Releases

Conditional Cognitive Biases in LLMs: How Biased User Turns Modulate In-Context Reasoning

DGX agent

arXiv:2608.05166v1 Announce Type: new Abstract: We present an evaluation of cognitive bias expression in state-of-the-art instruction-tuned LLMs under realistic multi-turn interaction settings. Our wo

model-releasesarxiv-cs-cl
7 Aug 2026
Safety

Decolonizing Linguistic Policies in Automated Speech Recognition: A Framework for Cross-Culturally Competent Speech AI

DGX agent

arXiv:2608.06141v1 Announce Type: new Abstract: This paper focuses on automatic speech recognition (ASR) and ASR-mediated voice interfaces that shape access to public services, healthcare, and educati

safetyarxiv-cs-cl
7 Aug 2026
Agents

Design and Evaluation of a Touchscreen-Based Teleoperation Interface for Robotic Manipulators

DGX agent

arXiv:2608.06219v1 Announce Type: new Abstract: Intuitive teleoperation interfaces are crucial for the safe and effective operation of robotic manipulators in challenging environments. In the nuclear

agentsarxiv-cs-ro
7 Aug 2026
Model Releases

EcoAgent-Bench: Evaluating Economic Decision-Making in Budget-Constrained LLM Agents

DGX agent

arXiv:2608.05519v1 Announce Type: new Abstract: Agent benchmarks usually measure task completion and treat resource use as an auxiliary statistic. In deployment, however, the choice among a local look

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

Estimating time spent on work tasks

DGX agent

arXiv:2608.05172v1 Announce Type: cross Abstract: The task-based framework in economics models occupations as bundles of tasks. It is the standard lens for understanding how technology affects work: a

safetyarxiv-cs-ai
7 Aug 2026
Safety

Faster and Better Alignment for Flow Matching Models via Step-aware Advantages

DGX agent

arXiv:2602.01591v2 Announce Type: replace Abstract: Recent advances in flow matching models, particularly with reinforcement learning (RL), have significantly enhanced human preference alignment in fe

safetyarxiv-cs-cv
7 Aug 2026
Safety

From Economic Agents to Agentic Economies: A Systems Blueprint for Economic World Models

DGX agent

arXiv:2608.06020v1 Announce Type: new Abstract: Economic World Models (EWMs) are generative economic models that simulate how economies evolve from within by modeling heterogeneous agents, their belie

safetyarxiv-cs-ai
7 Aug 2026
Safety

From Siloed Algorithms to Compliance-First Agentic Platforms: A Multi-Layered Architecture for Hospital AI Systems

DGX agent

arXiv:2608.06112v1 Announce Type: new Abstract: Hospitals are rapidly adopting artificial intelligence for triage, imaging, scheduling etc., yet most deployments remain isolated point solutions locked

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

GST-Bench: Can VLMs Develop Global Spatial Awareness from Video?

DGX agent

arXiv:2608.05747v1 Announce Type: new Abstract: Spatial intelligence is fundamental to embodied agents, yet existing benchmarks focus on local spatial perception from single or few viewpoints, overloo

model-releasesarxiv-cs-cv
7 Aug 2026
Safety

JTA: Joint Testability Architecture for Scenario-Based Validation of Safety-Critical Software

DGX agent

arXiv:2608.05594v1 Announce Type: cross Abstract: Validation adequacy in safety-critical software depends on more than the system under test. Critical scenarios must be constructed under controlled co

safetyarxiv-cs-ro
7 Aug 2026
Model Releases

MAC 2026: Advancing Micro-Action Analysis Towards Fine-Grained Understanding

DGX agent

arXiv:2607.16284v2 Announce Type: replace Abstract: Micro-Actions (MAs) are subtle and spontaneous human behaviors that provide important non-verbal cues in social interaction and affective communicat

model-releasesarxiv-cs-cv
7 Aug 2026
Safety

Mapping Patient-Perceived Physician Traits from Nationwide Online Reviews with LLMs

DGX agent

arXiv:2510.03997v2 Announce Type: replace Abstract: Understanding how patients perceive their physicians is essential to improving trust, communication, and satisfaction. Patients increasingly consult

safetyarxiv-cs-cl
7 Aug 2026
Model Releases

OmniMech: All-in-one Multimodal Mechanical Benchmark for 3D Reconstruction

DGX agent

arXiv:2608.05539v1 Announce Type: new Abstract: Recent vision-language models (VLMs) can generate executable CAD programs from images, but existing methods mainly target coarse, general-purpose 3D obj

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

PoolBench: A Benchmark for Pooling Strategies in Concept Representation Evaluation for Decoder-Only LLMs

DGX agent

arXiv:2608.05162v1 Announce Type: new Abstract: Pooling is a consequential but under-examined design choice in decoder-only concept representation work: practitioners must collapse token-level hidden

model-releasesarxiv-cs-cl
7 Aug 2026
Safety

Symbol Grounding in Neuro-Symbolic AI: A Gentle Introduction to Reasoning Shortcuts

DGX agent

arXiv:2510.14538v3 Announce Type: replace Abstract: Neuro-symbolic (NeSy) AI aims to develop deep neural networks whose predictions comply with prior knowledge encoding, e.g. safety or structural cons

safetyarxiv-cs-ai
7 Aug 2026
Agents

The Vulnerability With No CVE: Managing Persistent Gaps Between Mandate and Authority in AI Coding Agents

DGX agent

arXiv:2608.05884v1 Announce Type: cross Abstract: Existing guidance identifies excessive agency, excessive permission, weak task-bound authorization, and inadequate agent controls as important risks.

agentsarxiv-cs-cl
7 Aug 2026
← Previous
1…408409410411412…462
Next →