AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,797 results
Model Releases

Do Models Know Why They Changed Their Mind? Interpretability and Faithfulness of Chain-of-Thought Under Knowledge Conflict

DGX agent

arXiv:2605.27773v1 Announce Type: cross Abstract: When a language model sees a document contradicting its training knowledge, it must choose: follow the document or trust itself. Prior work proved thi

model-releasesarxiv-cs-ai
28 May 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Do We Really Need Quantum Machine Learning?: A Multidimensional Empirical Study

DGX agent

arXiv:2605.27923v1 Announce Type: cross Abstract: The rapid growth of computer vision and increasingly complex image recognition tasks has exposed fundamental computational limitations of classical ma

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Dr-CiK: A Testbed for Foresight-Driven Agents

DGX agent

arXiv:2605.27904v1 Announce Type: new Abstract: Time series forecasting in real-world settings often depends not only on historical observations, but also on external context that must be actively dis

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

DriveWAM: Video Generative Priors Enable Scalable World-Action Modeling for Autonomous Driving

DGX agent

arXiv:2605.28544v1 Announce Type: new Abstract: Pretrained foundation models have become an important basis for end-to-end autonomous driving. In contrast to vision-language models pretrained primaril

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

DRTriton: Large-Scale Synthetic Data Driven Reinforcement Learning for Triton Kernel Generation

DGX agent

arXiv:2603.21465v2 Announce Type: replace Abstract: Developing efficient CUDA kernels is a fundamental yet challenging task in the generative AI industry. Recent research leverages Large Language Mode

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

DynaSchedBench: Calibrated Dynamic Scheduling Benchmarks and Observability Paradox in LLM-based Scheduling Agents

DGX agent

arXiv:2605.27566v1 Announce Type: new Abstract: Progress in neural combinatorial optimization for Dynamic Flexible Job Shop Scheduling Problem (DFJSP) is currently hindered by a methodological tension

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Efficient and Scalable Provenance Tracking for LLM-Generated Code Snippets

DGX agent

arXiv:2605.28510v1 Announce Type: cross Abstract: Large language models (LLMs) for code completion and generation are increasingly used in software development, yet they may reproduce training example

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Efficient Pre-Training of LLMs through Truncated SVD Layers

DGX agent

arXiv:2605.28573v1 Announce Type: cross Abstract: The massive scaling of Large Language Models (LLMs) has made pretraining increasingly cost-prohibitive. While low-rank representation and orthonormal

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

EgoBench: An Interactive Egocentric Multimodal Benchmark for Tool-Using Agents

DGX agent

arXiv:2605.27820v1 Announce Type: new Abstract: As AI agents increasingly operate in open, real-world environments, they require a deep synergy of multimodal perception, tool invocation with multi-hop

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Energy-Structured Low-Rank Adaptation for Continual Learning

DGX agent

arXiv:2605.27482v1 Announce Type: cross Abstract: While orthogonal subspace methods try to mitigate task interference in Continual Learning (CL), they often suffer from energy diffusion across the bas

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Enhancing Trustworthy GUI Grounding via Self-Critiqued Reinforcement Learning

DGX agent

arXiv:2510.27266v2 Announce Type: replace Abstract: Autonomous graphical user interface (GUI) agents rely on accurate GUI grounding, which maps language instructions to on-screen coordinates, to execu

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

ESC-Skills: Discovering and Self-Evolving Skills for Emotional Support Conversations

DGX agent

arXiv:2605.27908v1 Announce Type: cross Abstract: Existing emotional support conversation (ESC) systems mainly rely on end-to-end response generation or coarse strategy supervision, offering limited i

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

EVADE-Bench: Multimodal Benchmark for Evaluating and Enhancing Evasive Content Detection

DGX agent

arXiv:2505.17654v4 Announce Type: replace-cross Abstract: E-commerce platforms increasingly rely on Large Language Models (LLMs) and Vision Language Models (VLMs) to detect illicit or misleading produ

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Evaluating Local Explainability Metrics for Machine Learning Models on Tabular Data

DGX agent

arXiv:2605.27618v1 Announce Type: new Abstract: Despite the wide use of explainability techniques to attempt to understand the behavior of Artificial Intelligence (AI), the generated explanations may

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

EventShiftFlow: Towards Hardware-efficient FPGA-based Flow Estimation

DGX agent

arXiv:2605.28312v1 Announce Type: cross Abstract: Event-based vision sensors offer asynchronous, high-temporal-resolution measurements that are attractive for low-latency robotic perception, but many

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Every9D-21M: Large-Scale Real-World 9D Canonicalization of Everyday Objects

DGX agent

arXiv:2605.28270v1 Announce Type: new Abstract: Estimating the 9D pose of everyday objects from a single real-world image remains challenging. This is largely due to the lack of large-scale supervisio

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Evolving Dataflow to process massive datasets for machine learning

DGX agent

Google created MapReduce more than 20 years ago to solve the scaling problems in data processing that the then young company was running into. The AI era that we are in now demands efficient, large-sc

model-releasesgoogle-cloud-ai
28 May 2026
Model Releases

EvoSpec: Evolving Speculative Decoding via Real-Time Vocabulary and Parameter AdaptationTarget

DGX agent

arXiv:2605.27390v1 Announce Type: cross Abstract: Speculative decoding accelerates Large Language Model inference via a draft-then-verify paradigm, yet the output projection layer becomes a bottleneck

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Exploratory Experience Shapes the Geometry of Predictive Representations

DGX agent

arXiv:2605.27929v1 Announce Type: cross Abstract: Active sensing links behavior and learning through an action-perception loop: actions determine the observations used to update internal predictive mo

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

FactReview: Evidence-Grounded Peer Review with Execution-Based Claim Verification

DGX agent

arXiv:2604.04074v3 Announce Type: replace Abstract: LLM-based reviewing systems typically take only the manuscript as input, leaving literature and code-based claims hard to verify. We present FactRev

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Federated Learning for Multivariate Time Series Anomaly Detection in Industrial Automation

DGX agent

arXiv:2605.27486v1 Announce Type: new Abstract: Federated learning (FL) has broadened the horizon for multivariate time series anomaly detection (MTSAD). However, benchmarking such anomaly detection m

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

FedMPT: Federated Multi-label Prompt Tuning of Vision-Language Models

DGX agent

arXiv:2605.28347v1 Announce Type: new Abstract: Multi-Label Recognition (MLR) based on Vision-Language Models (VLMs) aims to leverage their pre-trained knowledge to better adapt complex recognition sc

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Finding Miscompiles for Fun, Not Profit

DGX agent

This article likely discusses the discovery and analysis of compiler bugs or miscompilation errors in software development tools, exploring how developers identify these issues and their implications

model-releasessemianalysis
28 May 2026
Model Releases

Finding Miscompiles for Fun, Not Profit Or: You don’t need access to Claude Mythos to spend $10,000 in an afternoon https://newsletter.semia…

DGX agent

This post discusses how compiler bugs or miscompiles can be discovered and exploited without expensive AI systems, illustrating that significant computational costs can be incurred quickly when identi

model-releasesdylan-patel--x
28 May 2026
Model Releases

FLORO: A Multimodal Geospatial Foundation Model for Ecological Remote Sensing Across Sensors and Scales

DGX agent

arXiv:2605.28174v1 Announce Type: cross Abstract: Foundation models offer a promising route to transferable remote sensing representations, but many current approaches depend on very large pretraining

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

ForestHG-Trace: Traceable Long-Horizon Ecological Reasoning over Large-Scale Forest Scenes

DGX agent

arXiv:2605.27590v1 Announce Type: new Abstract: Remote sensing question answering (RS-QA) often requires more than direct semantic prediction, especially in large-scale forest scenes where ecological

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

FPMoE: A Sparse Mixture-of-Experts Approach to Functional Code Generation

DGX agent

arXiv:2605.27849v1 Announce Type: cross Abstract: Despite rapid progress in LLM-based code generation, existing models are predominantly trained on imperative languages, leaving functional programming

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Framing Matters: Addressing Framing Sensitivity in Decision-Making through Behaviorally-Grounded Value Alignment

DGX agent

arXiv:2605.28188v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in high-stakes decision-making settings such as legal reasoning, where consistency under factuall

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

From Detection to Mechanism: Cross-Attention Graph Neural Networks Enable Drug-Drug Interaction Type Prediction An Ablation Study with Acetylsalicylic Acid Validation

DGX agent

arXiv:2605.27861v1 Announce Type: cross Abstract: Predicting whether two drugs interact (binary detection) is a substantially dif- ferent task from predicting the mechanism type of that interaction (m

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

From Fact Overwriting to Knowledge Evolution: Causal Editing via On-Policy Self-Distillation

DGX agent

arXiv:2605.28303v1 Announce Type: new Abstract: While Knowledge Editing (KE) enables efficient updates, its dominant Static Fact Overwriting paradigm treats LLMs as discrete databases, forcibly inject

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

From Knowing to Doing: A Memory-Controlled Benchmark for LLM Trading Agents on Stock Markets

DGX agent

arXiv:2605.28359v1 Announce Type: new Abstract: Evaluating whether large language model (LLM) agents can profit in capital markets is increasingly framed as end-to-end trading: place an agent in a his

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

From paper to benchmark: agentic, framework-based reproduction of under-specified methods in machine health intelligence

DGX agent

arXiv:2605.28371v1 Announce Type: new Abstract: Industrial Prognostics and Health Management (PHM) provides a representative case study for a broader challenge in applied machine learning: translating

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Functional Entropy: Predicting Functional Correctness in LLM-Generated Code with Uncertainty Quantification

DGX agent

arXiv:2605.28500v1 Announce Type: cross Abstract: Large language models have shown impressive capabilities in code generation, yet they often produce functionally incorrect code. Uncertainty quantific

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players

DGX agent

arXiv:2605.28816v1 Announce Type: new Abstract: World models for interactive video generation have largely focused on single-agent settings, where future observations are generated from a single contr

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

glad to know Mythos' safety concerns have been addressed right as Anthropic also secured tens of billions in inference compute 👍

DGX agent

glad to know Mythos' safety concerns have been addressed right as Anthropic also secured tens of billions in inference compute 👍 JUST IN: Anthropic announces it will roll out Claude Mythos “in the com

model-releasesjeremy-howard--x
28 May 2026
Model Releases

Global Policy-Space Response Oracles for Two-Player Zero-Sum Games

DGX agent

arXiv:2605.28273v1 Announce Type: new Abstract: The Policy-Space Response Oracles (PSRO) framework scales equilibrium computation to large zero-sum games by iteratively expanding a restricted strategy

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Gradient Transformer: Learning to Generate Updates for LLMs

DGX agent

arXiv:2605.27591v1 Announce Type: new Abstract: Many organizations lack computational resources to fine-tune large language models (LLMs) on private (unshareable) data for better utility, while fine-t

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

GradientStabilizer:Fix the Norm, Not the Gradient

DGX agent

arXiv:2502.17055v4 Announce Type: replace-cross Abstract: Training instability in modern deep learning systems is frequently triggered by rare but extreme gradient-norm spikes, which can induce oversi

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Graph Neural Networks for Source Detection: A Review and Benchmark Study

DGX agent

arXiv:2512.20657v2 Announce Type: replace-cross Abstract: The source detection problem arises when an epidemic process unfolds over a contact network, and the objective is to identify its point of ori

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Graph-of-Skills: Dependency-Aware Structural Retrieval for Massive Agent Skills

DGX agent

arXiv:2604.05333v3 Announce Type: replace Abstract: Modern LLM agents increasingly rely on reusable skills, and as they interact with personal applications, web browsers, and other interfaces, skill l

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Hallucination Behavior in Multimodal LLMs Across Agricultural Image Interpretation and Generation Tasks

DGX agent

arXiv:2605.27595v1 Announce Type: cross Abstract: Large Language Models (LLMs) are being rapidly adopted in agricultural imaging applications, ranging from crop interpretation to synthetic field image

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

HardMTBench: Stress-Testing Chinese-English Translation on Knowledge-Intensive Domains

DGX agent

arXiv:2605.28315v1 Announce Type: new Abstract: General-purpose machine translation benchmarks such as FLORES-200 have reached a saturation regime on Chinese-English pairs, where modern large language

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows

DGX agent

arXiv:2605.27922v1 Announce Type: new Abstract: LLM agents are increasingly deployed as executable systems that use tools, modify workspaces, and produce concrete artifacts. In such workflows, perform

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

HELEA: Hard-Negative Benchmark and LLM-based Reranking for Robust Entity Alignment

DGX agent

arXiv:2605.28308v1 Announce Type: new Abstract: Entity Alignment (EA) is essential for knowledge graph (KG) fusion, but existing benchmarks often allow models to exploit name overlap rather than relat

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

hello beloved tasteful users, do you like how much claude thinks on your tasks? would love examples of it thinking too much or too little

DGX agent

This post asks Claude users for feedback about the length and depth of Claude's reasoning on tasks, soliciting examples where Claude's thinking might be excessive or insufficient. The inquiry aims to

model-releasesthariq--x
28 May 2026
Model Releases

Here Opus 4.8 built and play-tested a new RPG in Claude Code, including 3 PDF manuals and adventures, playtest notes, a website, and a playa…

DGX agent

Here Opus 4.8 built and play-tested a new RPG in Claude Code, including 3 PDF manuals and adventures, playtest notes, a website, and a playable solo adventure - then put it all on Netlify. No feedback

model-releasesethan-mollick--x
28 May 2026
Model Releases

High Performance, Low Reliability: Uncertainty Benchmarking for Tabular Foundation Models

DGX agent

arXiv:2605.28554v1 Announce Type: new Abstract: Recent Tabular Foundation Models (TFMs) have demonstrated state-of-the-art predictive performance, often surpassing Gradient-Boosted Decision Trees (GBD

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Holy smokes! Polymarket was not trolling. 500M accidental Claude spend in one month! Scoop from @MadisonMills22 @axios

DGX agent

Holy smokes! Polymarket was not trolling. 500M accidental Claude spend in one month! Scoop from @MadisonMills22 @axios NEW: AI consultant reveals a client accidentally spent $500,000,000.00 in a singl

model-releasesgary-marcus--x
28 May 2026
← Previous
1…253254255256257…475
Next →