AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
83,745 results
Research

Attn-QAT: 4-Bit Attention With Quantization-Aware Training

DGX agent

arXiv:2603.00040v3 Announce Type: replace-cross Abstract: Achieving reliable 4-bit attention is a prerequisite for end-to-end FP4 computation on emerging FP4-capable GPUs, yet attention remains the ma

researcharxiv-cs-ai
11 Aug 2026
Safety

Auditing Instruction-Trajectory Mismatches in Multimodal Robot Demonstrations

X Post
Paper
YouTube
Reddit
GitHub
DGX agent

arXiv:2608.07895v1 Announce Type: cross Abstract: Robot demonstration datasets used to train vision-language-action policies can contain a subtle but harmful failure mode: trajectories that are behavi

safetyarxiv-cs-lg
11 Aug 2026
Local Ai

Auditing Medical Vision-Language Models on Chest Radiographs: Estimating Reference Agreement Across Institutions

DGX agent

arXiv:2608.07550v1 Announce Type: new Abstract: Vision-language models return structured chest-radiograph findings through interfaces exposing no confidence score, so a receiving institution cannot re

local-aiarxiv-cs-cv
11 Aug 2026
Model Releases

Automated Generation of Complexity-Validated Decision Scenarios Using Large Language Models

DGX agent

arXiv:2608.08822v1 Announce Type: new Abstract: Cognitive decision-making research depends on diverse scenarios with carefully controlled complexity, yet manual production is slow, inconsistent, and b

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Automating Deception: Scalable Multi-Turn LLM Jailbreaks

DGX agent

arXiv:2511.19517v3 Announce Type: replace-cross Abstract: Multi-turn conversational attacks, which leverage psychological principles like Foot-in-the-Door (FITD), where a small initial request paves t

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Autonomous Driving with Priority-Ordered STL Specifications Under Multimodal Uncertainty

DGX agent

arXiv:2606.20336v2 Announce Type: replace Abstract: Autonomous vehicles must plan trajectories that satisfy multiple requirements, such as safety, traffic-rule compliance, and passenger comfort. Howev

safetyarxiv-cs-ro
11 Aug 2026
Safety

Autonomy Reshapes How Personalization Affects Privacy Concerns and Trust in LLM Agents

DGX agent

arXiv:2510.04465v3 Announce Type: replace-cross Abstract: LLM agents require personal information for personalization in order to effectively act on users' behalf, but this raises privacy concerns tha

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

AutoRefine: Compiling Trajectories into Validated Typed Agent Artifacts

DGX agent

arXiv:2601.22758v2 Announce Type: replace Abstract: Large language model agents repeatedly encounter related tasks, yet systems that learn from trajectories commit every lesson to one predefined artif

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Autorubric: A Unifying Framework for Rubric-Based LLM Evaluation on Non-Verifiable Tasks

DGX agent

arXiv:2603.00077v3 Announce Type: replace-cross Abstract: Rubric-based LLM judges have become indispensable for evaluating and optimizing systems on non-verifiable tasks, where success cannot be reduc

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Avalon-ToM-Bench: Evaluating Fine-Grained Theory of Mind via Asymmetric Game Mechanics

DGX agent

arXiv:2608.09638v1 Announce Type: new Abstract: Theory of Mind (ToM) is essential for agent interactions, yet existing evaluations either rely on static scenarios that oversimplify mental-state reason

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

b10356

DGX agent

ci : target ROCm 7.14 for build and release (#25775) Switch ROCm from 7.2.1 to 7.14 ROCm 7.14 is the first production release using TheRock build system. It can be installed using multi-arch deliverab

model-releasesllama-cpp-releases
11 Aug 2026
Model Releases

b10357

DGX agent

opencl: transpose the K tile in local memory for FA prefill kernels (#26428) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED ma

model-releasesllama-cpp-releases
11 Aug 2026
Model Releases

b10358

DGX agent

Address review comment of PR 25532 (#26852) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework L

model-releasesllama-cpp-releases
11 Aug 2026
Model Releases

b10359

DGX agent

ggml-webgpu: fix CI errors from #25025 and #25262 (#26566) test new flash_attn test rebase and fix to disable subgrou matrices when max_kv_tile == 0 delete log output Add i32 support to cpy and enable

model-releasesllama-cpp-releases
11 Aug 2026
Model Releases

b10360

DGX agent

common/peg : suppress incomplete escape sequences (#26780) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iO

model-releasesllama-cpp-releases
11 Aug 2026
Model Releases

b10361

DGX agent

model : fix SWA not being enabled for EXAONE 4.5 (#26848) model : fix SWA not being enabled for EXAONE 4.5 load_arch_hparams tests hparams.n_layer() == 64 before LLM_KV_NEXTN_PREDICT_LAYERS has been r

model-releasesllama-cpp-releases
11 Aug 2026
Model Releases

Back to the Future: A workbook time machine for spread sheet creation benchmarks

DGX agent

arXiv:2608.07873v1 Announce Type: new Abstract: We introduce the workbook time machine, a pipeline that automatically creates benchmarks evaluating the ability of language models to create derived obj

model-releasesarxiv-cs-ai
11 Aug 2026
Applications

Backward Compatibility in Tree-Based Explanations and Enhanced CART Algorithm

DGX agent

arXiv:2608.08674v1 Announce Type: new Abstract: In the operation of machine learning models, model update is a fundamental process that requires careful consideration of its impact on downstream decis

applicationsarxiv-cs-lg
11 Aug 2026
Model Releases

BAG: Budget-Aware Gating for Diffusion Caching

DGX agent

arXiv:2608.09231v1 Announce Type: new Abstract: Diffusion caching is a lightweight strategy that accelerates Diffusion Transformers (DiTs) by reusing intermediate features across denoising steps, but

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

BAP-MOS: Bandit-Based Adaptive Prompting for Boundary-Sensitive Multi-Organ Segmentation

DGX agent

arXiv:2608.08191v1 Announce Type: new Abstract: Multi-organ ultrasound segmentation remains challenging when anatomically adjacent structures must be delineated jointly, as localized boundary errors c

model-releasesarxiv-cs-cv
11 Aug 2026
Research

BASIS: Breach-Aware Selective Prompt Injection Shielding with Prefill Attention Probes

DGX agent

arXiv:2608.08027v1 Announce Type: cross Abstract: Prompt injection is a critical security threat in large language model (LLM) applications, where attackers hijack model behavior by embedding maliciou

researcharxiv-cs-lg
11 Aug 2026
Model Releases

Bayesian Symbolic Regression with Entropic Reinforcement Learning

DGX agent

arXiv:2608.09617v1 Announce Type: new Abstract: Symbolic regression is the problem of finding an algebraic expression describing a stochastic dependence of a target variable on a set of inputs. Unlike

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

BDH-CQ: In-Context Learning with Recurrent Latent Reasoning

DGX agent

arXiv:2608.09888v1 Announce Type: cross Abstract: We introduce BDH-CQ, a reasoning model that combines in-context learning with recurrent latent reasoning. Inputs presented at inference time continuou

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Benchmarking In-context Experiential Learning Through Repeated Product Recommendations

DGX agent

arXiv:2511.22130v2 Announce Type: replace Abstract: To navigate ever-shifting real-world environments, agents must grapple with incomplete knowledge and adapt their strategies through experience. Howe

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Benchmarking the Robustness of Agentic Systems to Adversarially-Induced Harms

DGX agent

arXiv:2508.16481v3 Announce Type: replace Abstract: Ensuring the safe use of agentic systems requires a thorough understanding of the range of malicious behaviors these systems may exhibit. In this pa

model-releasesarxiv-cs-lg
11 Aug 2026
Safety

Beyond Aggregate Calibration: Decomposing Income-Conditional Recall Disparities in Automated Credit Default Prediction

DGX agent

arXiv:2608.08202v1 Announce Type: new Abstract: Data-centric curation pipelines frequently rely on model confidence scores to flag and filter noisy or mislabeled training instances. Evaluating this fi

safetyarxiv-cs-lg
11 Aug 2026
Safety

Beyond Binary: Continuous State Optimization with Graph-Structured Objectives

DGX agent

arXiv:2608.09366v1 Announce Type: new Abstract: Large-scale learning systems often face the challenge of balancing multiple, potentially competing objectives, such as fairness, accuracy, and latency.

safetyarxiv-cs-lg
11 Aug 2026
Safety

Beyond cognacy

DGX agent

arXiv:2507.03005v3 Announce Type: replace Abstract: Computational phylogenetics has become an established tool in historical linguistics, with many language families now analyzed using likelihood-base

safetyarxiv-cs-cl
11 Aug 2026
Model Releases

Beyond Direct Identifiers: Probabilistic Privacy Risk Estimation for Privacy-Conscious LLM Query Delegation

DGX agent

arXiv:2608.09140v1 Announce Type: cross Abstract: Recent work on protecting privacy during user-LLM interactions often focuses on direct, explicit identifiers: the personally-identifiable information

model-releasesarxiv-cs-cl
11 Aug 2026
Research

Beyond Global Editing: Per-Instance Disentangled Subspaces for Training-Free Hallucination Mitigation in LVLMs

DGX agent

arXiv:2608.09344v1 Announce Type: new Abstract: Recent advances in large vision-language models (LVLMs) have enabled powerful multimodal reasoning by integrating visual encoders with large language mo

researcharxiv-cs-cv
11 Aug 2026
Local Ai

Beyond Hazard Resemblance: Contrastive Event Adjudication for Training-Free Video Anomaly Detection

DGX agent

arXiv:2608.09908v1 Announce Type: new Abstract: Video anomaly detection (VAD) aims to identify and temporally localize abnormal events in videos. Supervised methods learn anomaly decision boundaries f

local-aiarxiv-cs-cv
11 Aug 2026
Safety

Beyond 'I Can't Help With That': How Child Safety Experts Evaluate AI Chatbot Safety

DGX agent

arXiv:2608.07902v1 Announce Type: cross Abstract: Youth increasingly turn to AI chatbots for social and emotional support, raising concerns about how these systems respond, especially in high-stakes s

safetyarxiv-cs-ai
11 Aug 2026
Hardware

Beyond Isotropic Assumptions: Continuity-Constrained Segmentation and GPU Morphometry for Nanoscale GBM Analysis

DGX agent

arXiv:2608.07575v1 Announce Type: new Abstract: Confocal microscopy of optically cleared and swelled tissue resolves complex biological structures in 3D, but such acquisitions are highly anisotropic:

hardwarearxiv-cs-cv
11 Aug 2026
Model Releases

Beyond Naturalness: Probing Automated Text-To-Speech Evaluators on Linguistically Grounded Dimensions

DGX agent

arXiv:2608.09930v1 Announce Type: cross Abstract: Automated Text-to-Speech (TTS) evaluation methods (Mean Opinion Score (MOS) predictors and Audio Large Language Models (Audio-LLM) judges) are expecte

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Beyond Pixels: Benchmarking and Reward-Based Assessing Framework for Visual Spatial Aesthetics

DGX agent

arXiv:2512.05098v2 Announce Type: replace-cross Abstract: In recent years, Image Quality Assessment (IQA) for AI-generated images (AIGI) has advanced rapidly; however, existing methods primarily targe

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Beyond Pixels: Exploring DOM Downsampling for LLM-Based Web Agents

DGX agent

arXiv:2508.04412v3 Announce Type: replace Abstract: The advent of large language models (LLMs) has sparked an evolution of autonomous web browsing agents: given a web browsing task and serialised user

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Beyond Routing: Decoupling Expert Dispatch and Aggregation in Sparse Mixture-of-Experts

DGX agent

arXiv:2608.08853v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) routers commonly use the same scores both to select experts and to weight their already-computed outputs. We study wheth

model-releasesarxiv-cs-lg
11 Aug 2026
Safety

Beyond Solvability: Task Learnability as a Static Prior for LLM RL Post-Training

DGX agent

arXiv:2608.09217v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a central post-training paradigm for eliciting reasoning capabilities in large language models, yet uniform tas

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Beyond Static Models: An Evolving Framework for Continual Learning in Large Language Models across Training Stages

DGX agent

arXiv:2603.12658v2 Announce Type: replace-cross Abstract: Continual learning (CL) has emerged as a pivotal paradigm to enable large language models (LLMs) to dynamically adapt to evolving knowledge an

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Beyond Tables: Doc2DB-Bench for Relationally Faithful Document-to-Database Construction

DGX agent

arXiv:2608.08459v1 Announce Type: cross Abstract: Practical AI systems increasingly need to turn long, heterogeneous documents into queryable relational databases, not isolated spreadsheets. In domain

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Beyond the Capability Boundary: Zeroth-Order Optimization for Self-Evolving LLM Agents

DGX agent

arXiv:2608.09292v1 Announce Type: cross Abstract: Self-evolving methods improve the capabilities of LLM agents by sampling trajectories from the underlying LLMs and learning from these trajectories. H

model-releasesarxiv-cs-cl
11 Aug 2026
Tutorials

Beyond the Node: Clade-level Selection for Efficient MCTS in Automatic Heuristic Design

DGX agent

arXiv:2602.00549v2 Announce Type: replace Abstract: While Monte Carlo Tree Search (MCTS) shows promise in Large Language Model (LLM) based Automatic Heuristic Design (AHD), it suffers from a critical

tutorialsarxiv-cs-lg
11 Aug 2026
Local Ai

Beyond the Plane: Coupling Planar Vehicle Dynamics with Three-Dimensional Road Geometry

DGX agent

arXiv:2608.09402v1 Announce Type: new Abstract: Simulation is crucial for developing and testing autonomous driving systems. In particular, the development of localization and control algorithms relie

local-aiarxiv-cs-ro
11 Aug 2026
Local Ai

Beyond Uniform Restoration: Empowering All-in-One Restoration with Pixel-Level Multimodal Guidance

DGX agent

arXiv:2608.09482v1 Announce Type: cross Abstract: All-in-one image restoration is a unified low-level vision task that aims to effectively recover high-quality images from inputs degraded by various t

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

BibTeX Citation Errors in Scientific Publishing Agents: Evaluation and Mitigation

DGX agent

arXiv:2604.03159v2 Announce Type: replace-cross Abstract: Large language models with web search are increasingly used in scientific publishing agents, yet they produce BibTeX entries with pervasive fi

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents

DGX agent

arXiv:2608.09555v1 Announce Type: new Abstract: External natural-language skills provide large language model (LLM) agents with reusable and editable guidance for solving complex tasks. Yet their effe

model-releasesarxiv-cs-ai
11 Aug 2026
Hardware

Big Tech's AI boom echoes the 1870s railroad buildout, and Nvidia shifting risk to institutional capital may expose investors if AI revenues fail to materialize (Ben Thompson/Stratechery)

DGX agent

Ben Thompson / Stratechery: Big Tech's AI boom echoes the 1870s railroad buildout, and Nvidia shifting risk to institutional capital may expose investors if AI revenues fail to materialize — On Januar

hardwaretechmeme
11 Aug 2026
Model Releases

Biologically Informed Representation Learning for Robust Cross-Center Generalization of MALDI-TOF Mass Spectrometry

DGX agent

arXiv:2608.08182v1 Announce Type: cross Abstract: Machine learning models for MALDI-TOF mass spectrometry have shown considerable promise for clinical microbiology tasks such as microbial identificati

model-releasesarxiv-cs-ai
11 Aug 2026
← Previous
1…2829303132…1745
Next →