AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlog
90,259Total entries
1Added by human
90,258Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,128 results
Model Releases

What Do AI Agents Talk About? Discourse and Architectural Constraints in the First AI-Only Social Network

DGX agent

arXiv:2603.07880v5 Announce Type: replace Abstract: Moltbook is the first large-scale social network built for autonomous AI agent-to-agent interaction. Early studies on Moltbook have interpreted its

model-releasesarxiv-cs-cl
15 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

When Evidence Conflicts: Uncertainty and Order Effects in Retrieval-Augmented Biomedical Question Answering

DGX agent

arXiv:2605.14115v1 Announce Type: new Abstract: Biomedical retrieval-augmented large language models (LLMs) often face evidence that is incomplete, misleading, or internally contradictory, yet evaluat

researcharxiv-cs-cl
15 May 2026
Hardware

Woodelf++: A Fast and Unified Partial Dependence Plot Algorithm for Decision Tree Ensembles

DGX agent

arXiv:2605.14578v1 Announce Type: new Abstract: Partial Dependence Plots (PDPs) visualize how changes in a single feature affect the average model prediction. They are widely used in practice to inter

hardwarearxiv-cs-lg
15 May 2026
Safety

AgenticAITA: A Proof-Of-Concept About Deliberative Multi-Agent Reasoning for Autonomous Trading Systems

DGX agent

arXiv:2605.12532v1 Announce Type: cross Abstract: Conventional algorithmic trading systems are grounded in deterministic heuristics or offline-trained statistical models that cannot adapt to the seman

safetyarxiv-cs-ai
14 May 2026
Safety

Aligning Forest and Trees in Images & Long Captions for Visually Grounded Understanding

DGX agent

arXiv:2602.02977v2 Announce Type: replace-cross Abstract: Vision-language models such as CLIP often struggle to faithfully understand long, detail-rich captions, relying on dominant scene cues while o

safetyarxiv-cs-ai
14 May 2026
Research

ArcVQ-VAE: A Spherical Vector Quantization Framework with ArcCosine Additive Margin

DGX agent

arXiv:2605.13517v1 Announce Type: cross Abstract: Vector Quantized Variational Autoencoder (VQ-VAE) has become a fundamental framework for learning discrete representations in image modeling. However,

researcharxiv-cs-ai
14 May 2026
Safety

Asynchronous Reasoning: Training-Free Interactive Thinking LLMs

DGX agent

arXiv:2512.10931v3 Announce Type: replace Abstract: Many state-of-the-art LLMs are trained to think before giving their answer. Reasoning can greatly improve language model capabilities, but it also m

safetyarxiv-cs-lg
14 May 2026
Model Releases

Attention Once Is All You Need: Efficient Streaming Inference with Stateful Transformers

DGX agent

arXiv:2605.13784v1 Announce Type: new Abstract: Conventional transformer inference engines are request-driven, paying an O(n) prefill cost on every query. In streaming workloads, where data arrives co

model-releasesarxiv-cs-lg
14 May 2026
Safety

Auditing Sybil: Explaining Deep Lung Cancer Risk Prediction Through Generative Interventional Attributions

DGX agent

arXiv:2602.02560v2 Announce Type: replace-cross Abstract: Lung cancer remains the leading cause of cancer mortality, driving the development of automated screening tools to alleviate radiologist workl

safetyarxiv-cs-ai
14 May 2026
Research

Backdoor Channels Hidden in Latent Space: Cryptographic Undetectability in Modern Neural Networks

DGX agent

arXiv:2605.13214v1 Announce Type: cross Abstract: Recent cryptographic results establish that neural networks can be backdoored such that no efficient algorithm can distinguish them from a clean model

researcharxiv-cs-lg
14 May 2026
Research

Bayesian Nonparametric Mixed-Effect ODEs with Gaussian Processes

DGX agent

arXiv:2605.13088v1 Announce Type: new Abstract: Dynamical modelling is central to many scientific domains, including pharmacometrics, systems biology, physiology, and epidemiology. In these settings,

researcharxiv-cs-lg
14 May 2026
Research

Beyond Perplexity: A Geometric and Spectral Study of Low-Rank Pre-Training

DGX agent

arXiv:2605.13652v1 Announce Type: cross Abstract: Pre-training large language models is dominated by the memory cost of storing full-rank weights, gradients, and optimizer states. Low-rank pre-trainin

researcharxiv-cs-ai
14 May 2026
Research

Beyond Softmax: A Natural Parameterization for Categorical Random Variables

DGX agent

arXiv:2509.24728v2 Announce Type: replace Abstract: Latent categorical variables are frequently found in deep learning architectures. They can model actions in discrete reinforcement-learning environm

researcharxiv-cs-lg
14 May 2026
Research

BrainAnytime: Anatomy-Aware Cross-Modal Pretraining for Brain Image Analysis with Arbitrary Modality Availability

DGX agent

arXiv:2605.13059v1 Announce Type: new Abstract: Clinical diagnostic workups typically follow a modality escalation pathway: after initial clinical evaluation, clinicians begin with routine structural

researcharxiv-cs-cv
14 May 2026
Agents

Can LLM Agents Simulate Dynamic Networks? A Case Study on Email Networks with Phishing Synthesis

DGX agent

arXiv:2605.12507v1 Announce Type: cross Abstract: While Large Language Model (LLM) multi-agent systems (MAS) offer a transformative approach to simulating human behavior in complex systems, it remains

agentsarxiv-cs-ai
14 May 2026
Safety

Certified Robustness under Heterogeneous Perturbations via Hybrid Randomized Smoothing

DGX agent

arXiv:2605.12876v1 Announce Type: new Abstract: Randomized smoothing provides strong, model-agnostic robustness certificates, but existing guarantees are limited to single modalities, treating continu

safetyarxiv-cs-lg
14 May 2026
Model Releases

Cloud CISO Perspectives: How Google + Wiz changes multicloud strategy for CISOs

DGX agent

Welcome to the first Cloud CISO Perspectives for May 2026. Today, Vinod D’Souza, director, Office of the CISO, shares highlights from his RSA Conference fireside chat with Anthony Belfiore, chief stra

model-releasesgoogle-cloud-ai
14 May 2026
Model Releases

ConRetroBert: EMA Stabilized Dual Encoders for Template-Based Single-Step Retrosynthesis

DGX agent

arXiv:2605.12736v1 Announce Type: new Abstract: Template based single step retrosynthesis predicts reactants by selecting and applying an explicit reaction template, making each prediction traceable t

model-releasesarxiv-cs-lg
14 May 2026
Applications

Constraint-Aware Flow Matching: Decision Aligned End-to-End Training for Constrained Sampling

DGX agent

arXiv:2605.12754v1 Announce Type: new Abstract: Deep generative models provide state-of-the-art performance across a wide array of applications, with recent studies showing increasing applicability fo

applicationsarxiv-cs-lg
14 May 2026
Research

Context Training with Active Information Seeking

DGX agent

arXiv:2605.13050v1 Announce Type: cross Abstract: Most existing large language models (LLMs) are expensive to adapt after deployment, especially when a task requires newly produced information or nich

researcharxiv-cs-ai
14 May 2026
Model Releases

CoRe-Gen: Robust Spectrum-to-Structure Generation under Imperfect Fingerprint Conditions

DGX agent

arXiv:2605.12980v1 Announce Type: cross Abstract: Molecular structure elucidation from tandem mass spectra (MS/MS) remains challenging, particularly for de novo generation beyond database coverage. A

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

CUBic: Coordinated Unified Bimanual Perception and Control Framework

DGX agent

arXiv:2605.13452v1 Announce Type: cross Abstract: Recent advances in visuomotor policy learning have enabled robots to perform control directly from visual inputs. Yet, extending such end-to-end learn

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Deepseek V4 Flash is now free via Nous Portal for a limited time thanks to @novita_labs!

DGX agent

Nous Research announced that Deepseek V4 Flash is temporarily available for free access through the Nous Portal, courtesy of Nous Research in collaboration with @novita_labs. This limited-time offer p

model-releasesnous-research--x
14 May 2026
Safety

Digital Twins as Synthetic Controls in Single-Arm Trials

DGX agent

arXiv:2605.12832v1 Announce Type: cross Abstract: Single-arm trials are an important study design for evaluating drug efficacy and safety without enrolling patients into a control arm. Although they d

safetyarxiv-cs-lg
14 May 2026
Research

DirectTryOn: One-Step Virtual Try-On via Straightened Conditional Transport

DGX agent

arXiv:2605.12939v1 Announce Type: new Abstract: Recent diffusion- and flow-based VTON methods achieve strong results with pretrained generative models, but their reliance on multi-step sampling incurs

researcharxiv-cs-cv
14 May 2026
Tutorials

Distribution Shift in Missing Data Imputation: A Risk-Based Perspective and Importance-Weighted Correction under MAR

DGX agent

arXiv:2602.06713v2 Announce Type: replace-cross Abstract: Missing data imputation, where a model is trained on observed data to estimate unobserved values, is a fundamental problem in machine learning

tutorialsarxiv-cs-lg
14 May 2026
Applications

Do Heavy Tails Help Diffusion? On the Subtle Trade-off Between Initialization and Training

DGX agent

arXiv:2605.13175v1 Announce Type: new Abstract: Recent works have proposed incorporating heavy-tailed (HT) noise into diffusion- and flow-based generative models, with the goals of better recovering t

applicationsarxiv-cs-lg
14 May 2026
Research

Does language matter for spoken word classification? A multilingual generative meta-learning approach

DGX agent

arXiv:2605.13084v1 Announce Type: cross Abstract: Meta-learning has been shown to have better performance than supervised learning for few-shot monolingual spoken word classification. However, the met

researcharxiv-cs-ai
14 May 2026
Research

Early Data Exposure Improves Robustness to Subsequent Fine-Tuning

DGX agent

arXiv:2605.12705v1 Announce Type: new Abstract: How can we train models whose post-trained capabilities survive subsequent fine-tuning? Rather than focusing on downstream interventions to mitigate for

researcharxiv-cs-lg
14 May 2026
Safety

Emergent and Subliminal Misalignment Through the Lens of Data-Mediated Transfer

DGX agent

arXiv:2605.12798v1 Announce Type: cross Abstract: Fine-tuning LLMs on narrow harmful datasets can induce Emergent Misalignment (EM), where models exhibit misaligned behavior far beyond the fine-tuning

safetyarxiv-cs-ai
14 May 2026
Model Releases

Exact Sequence Interpolation with Transformers

DGX agent

arXiv:2502.02270v3 Announce Type: replace Abstract: We prove that transformers can exactly interpolate datasets of finite input sequences in R^d, dgeq 2, with corresponding output sequences of smaller

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Fireworks Training Platform continues to expand. Today GLM 5.1 LoRA RL is now live via Training API: SFT, DPO, and full RL on a 200K context…

DGX agent

Fireworks Training Platform continues to expand. Today GLM 5.1 LoRA RL is now live via Training API: SFT, DPO, and full RL on a 200K context window → custom loss functions or smart defaults. No usage

model-releasesfireworks-ai--x
14 May 2026
Model Releases

Force-Aware Neural Tangent Kernels for Scalable and Robust Active Learning of MLIPs

DGX agent

arXiv:2605.13788v1 Announce Type: new Abstract: Active learning for machine-learning interatomic potentials (MLIPs) must address several challenges to be practical: scaling to large candidate pools, l

model-releasesarxiv-cs-lg
14 May 2026
Research

From Baselines to Transport Geodesics: Axiomatic Attribution via Optimal Generative Flows

DGX agent

arXiv:2603.05093v2 Announce Type: replace-cross Abstract: Feature attributions often hide a critical modeling choice: they explain a prediction along a counterfactual path from a reference state to an

researcharxiv-cs-ai
14 May 2026
Research

From Generalist to Specialist Representation

DGX agent

arXiv:2605.12733v1 Announce Type: cross Abstract: Given a generalist model, learning a task-relevant specialist representation is fundamental for downstream applications. Identifiability, the asymptot

researcharxiv-cs-ai
14 May 2026
Research

GeoFlowVLM: Geometry-Aware Joint Uncertainty for Frozen Vision-Language Embedding

DGX agent

arXiv:2605.13352v1 Announce Type: new Abstract: Standard dual-encoder vision-language models that map images and text to deterministic points on a shared unit hypersphere through ell_2 normalization t

researcharxiv-cs-lg
14 May 2026
Model Releases

Geometric Preconditioning and Curriculum Optimization for Trainable Variational Quantum Regression

DGX agent

arXiv:2601.11942v3 Announce Type: replace Abstract: Variational quantum circuits are increasingly studied as continuous-function approximators, but quantum regression remains difficult to train when g

model-releasesarxiv-cs-lg
14 May 2026
Local Ai

GLASS: Global-Local Aggregation for Inference-time Sparsification of LLMs

DGX agent

arXiv:2508.14302v2 Announce Type: replace-cross Abstract: Inference-time sparsification is a promising path to deploy large language models (LLMs) on resource-constrained devices, yet existing trainin

local-aiarxiv-cs-ai
14 May 2026
Tutorials

GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking

DGX agent

arXiv:2602.17555v3 Announce Type: replace Abstract: Video reasoning requires a fine-grained understanding of the temporal dependencies and event-level relations between objects and events in videos. C

tutorialsarxiv-cs-cv
14 May 2026
Model Releases

Hierarchical Transformer Preconditioning for Interactive Physics Simulation

DGX agent

arXiv:2605.13343v1 Announce Type: cross Abstract: Neural preconditioners for real-time physics simulation offer promising data-driven priors, but they often fail to capture long-range couplings effici

model-releasesarxiv-cs-lg
14 May 2026
Safety

History Anchors: How Prior Behavior Steers LLM Decisions Toward Unsafe Actions

DGX agent

arXiv:2605.13825v1 Announce Type: new Abstract: Frontier LLMs are increasingly deployed as agents that pick the next action after a long log of prior tool calls produced by the same or a different mod

safetyarxiv-cs-ai
14 May 2026
Model Releases

Implicit Behavioral Decoding from Next-Step Spike Forecasts at Population Scale

DGX agent

arXiv:2605.12999v1 Announce Type: cross Abstract: Closed-loop brain-computer interfaces often require both a forecast of upcoming neural population activity and a readout of the animal's behavioral st

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Imposing Boundary Conditions on Neural Operators via Learned Function Extensions

DGX agent

arXiv:2602.04923v2 Announce Type: replace Abstract: Neural operators have emerged as powerful surrogates for the solution of partial differential equations (PDEs), yet their ability to handle general,

model-releasesarxiv-cs-lg
14 May 2026
Safety

Improving Classifier-Free Guidance of Flow Matching via Manifold Projection

DGX agent

arXiv:2601.21892v2 Announce Type: replace-cross Abstract: Classifier-free guidance (CFG) is a widely used technique for controllable generation in diffusion and flow-based models. Despite its empirica

safetyarxiv-cs-ai
14 May 2026
Model Releases

IndicMedDialog: A Parallel Multi-Turn Medical Dialogue Dataset for Accessible Healthcare in Indic Languages

DGX agent

arXiv:2605.13292v1 Announce Type: cross Abstract: Most existing medical dialogue systems operate in a single-turn question--answering paradigm or rely on template-based datasets, limiting conversation

model-releasesarxiv-cs-ai
14 May 2026
Research

Information as Maximum-Caliber Deviation: A bridge between Integrated Information Theory and the Free Energy Principle

DGX agent

arXiv:2605.12536v1 Announce Type: cross Abstract: The Free Energy Principle (FEP) is a leading framework for mathematically modeling self-organization and learning, while Integrated Information Theory

researcharxiv-cs-ai
14 May 2026
Safety

Interesting position paper on agentic AI as a foreseeable pathway to AGI. (bookmark it) There has been strong debate on whether a larger sin…

DGX agent

Interesting position paper on agentic AI as a foreseeable pathway to AGI. (bookmark it) There has been strong debate on whether a larger single model get us there or a multi-agent system. The authors

safetydair-ai--x
14 May 2026
Model Releases

Learning Responsibility-Attributed Adversarial Scenarios for Testing Autonomous Vehicles

DGX agent

arXiv:2605.13751v1 Announce Type: new Abstract: Establishing trustworthy safety assurance for autonomous driving systems (ADSs) requires evidence that failures arise from avoidable system deficiencies

model-releasesarxiv-cs-ro
14 May 2026
← Previous
1…716717718719720…1357
Next →