AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

88,343Total entries
1Added by human
88,342Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,572 results
19 Aug 2026

TraceSQL: Traceable Answerability Estimation for Reference-Free Text-to-SQL Verification

SafetyDGX agent

arXiv:2608.17795v1 Announce Type: new Abstract: Text-to-SQL systems are commonly evaluated using ground-truth SQL queries or reference execution results, but such supervision is unavailable at inferen

VLCP: Vision Language Control Policy Closed-Loop Code Replanning for Robot Manipulation

SafetyDGX agent

arXiv:2608.16978v1 Announce Type: new Abstract: Turning a frontier vision-language model into a robot policy usually means fine-tuning it to emit an action representation it never saw in pretraining,

what is happening here: claude used existing open source tools and orchestrated them together to do a protein design campaign (with a big gu…

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

what is happening here: claude used existing open source tools and orchestrated them together to do a protein design campaign (with a big guide prompt!) this is a win for claude because doing this orc

Whether LLMs Can Navigate Beliefs and Facts Depends on How You Phrase It

ResearchDGX agent

arXiv:2608.17809v1 Announce Type: new Abstract: Humans naturally form and express beliefs in daily communication, e.g., 'I think the answer is 3' or 'I suppose that's right.' Such beliefs inevitably i

Wuying-Browser-Agent: Real-World Centric Fundamental Long-Horizon Browser Agents

Model ReleasesDGX agent

arXiv:2608.17319v1 Announce Type: new Abstract: Browser agents perform well on short, clean demonstrations, but real deployment is fundamentally different: agents must sustain dozens of decisions on l

18 Aug 2026

A Human-Centred Approach to Benchmarking LLMs for Parenting Advice

ResearchDGX agent

arXiv:2608.14622v1 Announce Type: new Abstract: People are increasingly using large language models (LLMs) to seek advice, including for parenting. Parenting is a critical and socially sensitive domai

A Machine-Learned Comorbidity Index

Model ReleasesDGX agent

arXiv:2606.17450v2 Announce Type: replace Abstract: Traditional comorbidity scores (e.g., Charlson and Elixhauser) are widely used for risk adjustment and patient stratification, but they have two key

A Parameter-Free Few-Shot Evaluation for Elephant Vocalisation Classification

Model ReleasesDGX agent

arXiv:2608.14824v1 Announce Type: cross Abstract: We present a parameter-free episodic evaluation of nearest-centroid classification for elephant vocalisations on fixed pretrained acoustic embeddings,

A Two-Stage Learning PINN Approach for Solving the Inverse Problem of the 1D Porous Medium Equation

Model ReleasesDGX agent

arXiv:2608.16475v1 Announce Type: cross Abstract: The Porous Medium Equation (PME), given by u_t = Delta(u^m) for m > 1, is a degenerate nonlinear parabolic partial differential equation that arises i

AdROD: HyperNetwork-based Adversarially Robust Object Detection for Autonomous Driving

Model ReleasesDGX agent

arXiv:2608.16031v1 Announce Type: cross Abstract: Camera-based object detectors are vulnerable to physical adversarial attacks designed to suppress detections. While adversarial training and input pur

Agentic-SQL Revisited: Autonomy-Based Taxonomy and Empirical Benchmark Analysis for LLM Text-to-SQL

Model ReleasesDGX agent

arXiv:2608.15389v1 Announce Type: new Abstract: LLM-based Text-to-SQL progress is reported across heterogeneous benchmarks, backbones, and inference protocols, making cross-system comparison fragile.

Algorithm-Architecture Co-Design for Efficient VLA Inference via Speculative Inference and Verification

HardwareDGX agent

arXiv:2608.15636v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable capabilities in the field of embodied AI, but their high computational cost and limit

Amortised Post-Hoc Explanation with Exact Preservation for Dynamic Graph Anomaly Detectors

Model ReleasesDGX agent

arXiv:2608.15559v1 Announce Type: cross Abstract: Anomaly detection in dynamic graphs underpins financial fraud analysis, intrusion detection, and platform integrity, where automated decisions require

An Analytical-Prior Framework for Data-Efficient Prediction of Sound-Reduction Frequencies in Rectangular Side-Branch Helmholtz Resonators

TutorialsDGX agent

arXiv:2608.16873v1 Announce Type: new Abstract: High-fidelity finite-element simulations can provide accurate numerical predictions for side-branch resonators, but large simulation datasets are expens

AnchorScore: A CLIP-Based Diagnostic of MLLM Annotation Difficulty

ResearchDGX agent

arXiv:2608.16690v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are widely used for automated annotation, yet their per-class accuracy varies widely (e.g., 12%-98% across the

Anthropic details two experiments showing how Claude can accelerate protein design and analytical chemistry, and says it plans an access program for scientists (Anthropic)

Model ReleasesDGX agent

Anthropic: Anthropic details two experiments showing how Claude can accelerate protein design and analytical chemistry, and says it plans an access program for scientists — Summary: In this post, we s

ARGUS: Attention-Guided Transformers for Scalable Person Identification Using Wi-Fi Telemetry

Model ReleasesDGX agent

arXiv:2608.14670v1 Announce Type: cross Abstract: Passive, device-free person identification offers an alternative to camera- and wearable-based biometrics, yet existing wireless approaches rely large

Assessing LLMs' mathematical abilities requires understanding the various mechanisms of mathematical creativity

ApplicationsDGX agent

arXiv:2608.16118v1 Announce Type: new Abstract: How should we assess whether large language models can perform mathematical invention? I argue that this question is currently underspecified: mathemati

AsyTO: Asymmetric Temporal Operator for Parameter-Efficient Multivariate Time Series Forecasting

Model ReleasesDGX agent

arXiv:2608.16098v1 Announce Type: cross Abstract: Multivariate time-series forecasting faces a structural dilemma: sharing one temporal predictor across variables is parameter-efficient but forces het

AudioTQ: A Data-Oblivious 6-Bit CPU Audio Codec via Randomized Hadamard Rotation and Lloyd-Max Quantization

ResearchDGX agent

arXiv:2608.15369v1 Announce Type: cross Abstract: Lossy audio compression algorithms traditionally rely on psychoacoustic modeling and frequency-domain representations (e.g., MP3, AAC, and Opus) to di

AutoSR: Automatic Symbolic Regression by Searching Research States

Model ReleasesDGX agent

arXiv:2608.16876v1 Announce Type: cross Abstract: We introduce Automatic Symbolic Regression (AutoSR), a fully automated system that instantiates Research-Space Symbolic Regression by searching persis

b10483

Model ReleasesDGX agent

build : fix xcframework + cmake clean-up (#27304) xcframework : fix build mtmd : remove unused include path vendor : use vendor::hash alias target in cmake CMake reserves '::' in target names for impo

b10485

Model ReleasesDGX agent

sync : ggml Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubuntu x64 (CPU) Ubuntu ar

b10486

Model ReleasesDGX agent

mtmd: fix LFM2 image tiling threshold (#27057) mtmd: fix LFM2 image tiling threshold refactor testing fix fix on windows Co-authored-by: Xuan Son Nguyen son@huggingface.co Website: https://llama.app m

b10488

Model ReleasesDGX agent

ci : Update OpenVINO to 2026.3, skip nemotron-h rollback test (#27292) update to ov-2026.3, update device drivers ci: skip nemotron-h rollback test on OpenVINO The OpenVINO backend does not support SS

BaT: Towards Self-Evolving Medical Research Agent with Stage Rubrics

Model ReleasesDGX agent

arXiv:2608.16211v1 Announce Type: new Abstract: Long-horizon agents are beginning to automate complete workflows that produce code, reports, and research artifacts. Medical imaging workflows are multi

Behaviour Is an Incomplete Measure of Reasoning Development: Cross-surface pre-arrival accessibility and the limits of developmental inference in a recurrent-depth reasoner

Model ReleasesDGX agent

arXiv:2608.16085v1 Announce Type: cross Abstract: Capability development is routinely inferred from behavioural thresholds, from final checkpoints, or from what a decoder can read out of a hidden stat

Beyond Correctness: Toward Automated Novelty Verification with Lean 4

Model ReleasesDGX agent

arXiv:2608.14669v1 Announce Type: new Abstract: Artificial intelligence systems applied to mathematics verify correctness but not novelty: an automatically generated theorem can compile in Lean withou

Beyond Direct Access: Resource Hijacking in LLM Agents

Local AiDGX agent

arXiv:2608.15108v1 Announce Type: cross Abstract: Large language model agents are increasingly connected to high-value resources such as computing infrastructure, credentials, usage budgets, identitie

Beyond Field Accuracy: Two-Axis Diagnosis of Inverse-PINN Parameter Error

Model ReleasesDGX agent

arXiv:2608.15373v1 Announce Type: new Abstract: Inverse physics-informed neural networks (PINNs) can reconstruct a field accurately while returning an incorrect physical parameter. We introduce a two-

Beyond Pass@k: Measuring Reliability and Security of Agentic Code Generation

Model ReleasesDGX agent

arXiv:2608.14711v1 Announce Type: new Abstract: AI coding agent benchmarks rank agents with the Chen et al. (2021) pass@k estimator, but current implementations misapply it: they set n to the number o

Beyond Similarity Matching: Structured Reasoning for Open-Vocabulary Referring Segmentation in 3DGS

SafetyDGX agent

arXiv:2608.16103v1 Announce Type: new Abstract: Open-vocabulary referring segmentation in 3D Gaussian Splatting (3DGS) requires a neural model to select Gaussian primitives according to free-form lang

Bounded Agents: Delegation Security for Multi-Agent AI Systems

AgentsDGX agent

arXiv:2608.15888v1 Announce Type: new Abstract: LLM-based agents can act on behalf of a user to access cloud services, call tools, or invoke agents. At session start, the agent's permissions are set b

Can Attribution Predict Risk? From Multi-View Attribution to Planning Risk Signals in End-to-End Autonomous Driving

AgentsDGX agent

arXiv:2605.06264v2 Announce Type: replace Abstract: End-to-end autonomous driving models generate future trajectories from multi-view inputs, improving system integration but introducing opaque decisi

CAPO: Constraint-Aware Prompt Optimization for LLM Agents

SafetyDGX agent

arXiv:2608.16068v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as agents that rely on system prompts to use tools and complete tasks. Such deployments impose

CFR without Unbiasedness: Deterministic Guarantees for Persistent Public-Chance Schedules

Model ReleasesDGX agent

arXiv:2608.14761v1 Announce Type: cross Abstract: At a finite public-chance cut, counterfactual regret minimization (CFR) must choose how many outcomes to evaluate before each regret update. Exact eva

CompoSkill: Compositional Skill Chain Attacks from Individually Scanner-Passing LLM Agent Skills

Model ReleasesDGX agent

arXiv:2608.16246v1 Announce Type: cross Abstract: Autonomous AI agents tackling Long Horizon Tasks depend on marketplace skills that are certified one at a time: a scanner returns a safety verdict for

Comprehension of Multilingual Expressions Referring to Target Objects in Visual Inputs

Local AiDGX agent

arXiv:2511.11427v2 Announce Type: replace Abstract: Referring Expression Comprehension (REC) requires models to localize objects in images based on different types of natural language descriptions. Ev

ContextClaim: A Context-Driven Paradigm for Verifiable Claim Detection

ResearchDGX agent

arXiv:2603.30025v2 Announce Type: replace Abstract: Automated fact-checking pipelines typically begin with a filtering stage that decides which claims are worth verifying, given that the later evidenc

Continuous Quantum Feedback Control via Kraus-Parameterized Belief Reinforcement Learning

Model ReleasesDGX agent

arXiv:2608.15715v1 Announce Type: cross Abstract: Quantum feedback control requires acting on noisy continuous measurement records without direct access to the underlying quantum state. We propose Kra

Convolution-Free Holistic Multivariance Decomposition Layer for Efficient Hyperspectral Image Classification Tensor Networks

Model ReleasesDGX agent

arXiv:2608.16241v1 Announce Type: new Abstract: Feature extraction for hyperspectral image classification is conventionally addressed using rigid tensor decompositions that fail to capture complex spa

Convolution Smoothed Quantile Regression for XGBoost

Model ReleasesDGX agent

arXiv:2608.15290v1 Announce Type: cross Abstract: The increasing availability of large and complex datasets across many scientific disciplines has led to widespread adoption of machine learning (ML) f

CrevasseSeg: A Label-Efficient UAV Crevasse Segmentation Framework

Model ReleasesDGX agent

arXiv:2608.15790v1 Announce Type: new Abstract: Crevasse mapping from uncrewed aerial vehicle (UAV) imagery matters for glaciological research and for field safety in glaciated terrain. Yet, pixel-lev

D^{2}R^{2}: Discrete Diffusion with Regulation Reinforcement for Single-Cell Perturbation Prediction

SafetyDGX agent

arXiv:2608.15288v1 Announce Type: new Abstract: Predicting single-cell transcriptomic responses to genetic perturbations is central to functional genomics and virtual-cell modeling. Existing approache

Decentralized Federated Learning for Heterogeneous Multi-Task Semantic Communication

Model ReleasesDGX agent

arXiv:2608.15256v1 Announce Type: new Abstract: Collaborative training in distributed semantic communication (DSC) networks typically relies on decentralized federated learning (DFL). However, pushing

Deep Thought Alignment: Trajectory-Level Latent Distillation for Video Reasoning

SafetyDGX agent

arXiv:2608.16316v1 Announce Type: cross Abstract: Large Multimodal Models (LMMs) for video reasoning have long been hindered by the high computational cost of processing vast amounts of visual informa

DeepOHeat-v2: Self-Improving Operator Learning for Fast and Trustworthy Thermal Optimization in 3D-IC Design

Model ReleasesDGX agent

arXiv:2608.16080v1 Announce Type: new Abstract: Thermal-aware optimization of multi-die 3D integrated circuits evaluates many designs, each a costly heat-equation solve. Operator-learning surrogates r

DeepSeek V4 Pro 0813 vs GPT-5.6 Sol on DeepSWE: Cost, Coding, and Routing

Model ReleasesDGX agent

In the DeepSWE benchmark a two‑stage approach of running DeepSeek V4 Pro 0813 first and falling back to GPT‑5.6 Sol when the tests fail achieves 83 % success at an average cost of 3.35 per task, compa

Degradation-Aligned Self-Supervised Learning for State of Health Estimation of Lithium-Ion Batteries under Label Sparsity

ApplicationsDGX agent

arXiv:2608.16612v1 Announce Type: cross Abstract: An accurate estimation of the state of health (SOH) underpins a safe and optimized use of the battery system. Although compelling, data-driven SOH est

Depth-guided Multi-view Exposure Bracketing for HDR Robot Vision

Model ReleasesDGX agent

arXiv:2608.16014v1 Announce Type: new Abstract: Achieving reliable single-shot high dynamic range (HDR) imaging under extreme illumination conditions remains a long-standing challenge, yet no comprehe

Detecting Money Laundering in Rwandan Mobile Money: A Machine Learning Framework

Model ReleasesDGX agent

arXiv:2608.15447v1 Announce Type: new Abstract: Mobile money has widened financial access across Sub-Saharan Africa and enlarged the surface for money-laundering and terrorism-financing (ML/TF) activi

Do Geometry-Aware Positional Encodings Help Transformers in Spatial Imperfect-Information Games?

Model ReleasesDGX agent

arXiv:2608.14982v1 Announce Type: cross Abstract: Transformers applied to spatial imperfect-information games must represent map geometry while tracking hidden entities through time. We ask whether ge

Does an Ollama-ready build keep this 27B's MTP and vision stack?

Model ReleasesDGX agent

One line in a new Qwen3.8-27B derivative card caught my eye: the uploader says the vision-language tower and MTP head were preserved through its offline block-FP8 build. The checkpoint is OrcaRouter's

Does generative AI supersede supervised XMLC? A Benchmark Study on Automated Subject Indexing with German Scientific Literature

Model ReleasesDGX agent

arXiv:2607.14882v2 Announce Type: replace-cross Abstract: With a large controlled vocabulary as the label set, the task of automated subject indexing in a library can be understood as a multi-label cl

DRAFE: Domain-Robust Asymmetric Fusion of Heterogeneous Detection Transformers for Cross-City Fine-Grained Traffic Object Detection

Model ReleasesDGX agent

arXiv:2608.16632v1 Announce Type: new Abstract: Deep learning-based object detectors are fundamental to intelligent transportation systems, enabling traffic monitoring, vehicle analytics, and infrastr

Drive, Pack, Fly: The Travelling Thief Problem with Drone

Model ReleasesDGX agent

arXiv:2608.16435v1 Announce Type: new Abstract: In collection operations, accumulating payload progressively slows the vehicle, imposing a cumulative penalty on routing efficiency. An onboard drone ca

EdgeCrafter: Compact ViTs for Edge Dense Prediction via Task-Specialized Distillation

ResearchDGX agent

arXiv:2603.18739v4 Announce Type: replace Abstract: Deploying high-performance dense prediction models on resource-constrained edge devices remains challenging due to strict computation and memory bud

Efficient Coreset Selection via K-Nearest Neighbor Graphs

ApplicationsDGX agent

arXiv:2608.16270v1 Announce Type: new Abstract: Coreset selection reduces the cost of model training by replacing a large training set with a small representative subset. Existing gradient-approximati

Emergent Misaligned Communication in Long-Horizon Multi-Agent LLM Commerce

SafetyDGX agent

arXiv:2608.14825v1 Announce Type: cross Abstract: Frontier LLM agents increasingly transact on behalf of separate principals, often using natural language rather than structured APIs. Much of the safe

Empowering Polymeric Materials Discovery by Artificial Intelligence

AgentsDGX agent

arXiv:2606.20753v2 Announce Type: replace-cross Abstract: Polymeric materials underpin modern technologies spanning energy storage, microelectronics, healthcare and sustainable manufacturing. Yet thei

← Previous
1…593594595596597…1060
Next →