AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,624 results
15 May 2026

Seeing Through the Brain: New Insights from Decoding Visual Stimuli with fMRI

ApplicationsDGX agent

arXiv:2510.16196v2 Announce Type: replace-cross Abstract: Understanding how the brain encodes visual information is a central challenge in neuroscience and machine learning. A promising approach is to

SpeakerLLM: A Speaker-Specialized Audio-LLM for Speaker Understanding and Verification Reasoning

SafetyDGX agent

arXiv:2605.15044v1 Announce Type: cross Abstract: As audio-first agents become increasingly common in physical AI, conversational robots, and screenless wearables, audio large language models (audio-L

SToRe3D: Sparse Token Relevance in ViTs for Efficient Multi-View 3D Object Detection

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.14110v1 Announce Type: new Abstract: Vision Transformers (ViTs) enable strong multi-view 3D detection but are limited by high inference latency from dense token and query processing across

StyleTextGen: Style-Conditioned Multilingual Scene Text Generation

Model ReleasesDGX agent

arXiv:2605.14708v1 Announce Type: new Abstract: Style-conditioned scene text generation faces unique challenges in extracting precise text styles from complex backgrounds and maintaining fine-grained

SVAG-Bench: A Large-Scale Benchmark for Multi-Instance Spatio-temporal Video Action Grounding

Model ReleasesDGX agent

arXiv:2510.13016v3 Announce Type: replace Abstract: A truly capable AI system must do more than detect objects or recognize activities in isolation. It must form unified, grounded representations of w

Sycophancy is an Educational Safety Risk: Why LLM Tutors Need Sycophancy Benchmarks

Model ReleasesDGX agent

arXiv:2605.14604v1 Announce Type: new Abstract: This position paper argues that effective tutoring requires corrective friction: surfacing misconceptions and challenging them supportively to drive con

The Great Pretender: A Stochasticity Problem in LLM Jailbreak

SafetyDGX agent

arXiv:2605.14418v1 Announce Type: cross Abstract: 'Oh-Oh, yes, I'm the great pretender. Pretending that I'm doing well. My need is such, I pretend too much...' summarizes the state in the area of jail

To discretize continually: Mean shift interacting particle systems for Bayesian inference

Model ReleasesDGX agent

arXiv:2605.14142v1 Announce Type: cross Abstract: Integration against a probability distribution given its unnormalized density is a central task in Bayesian inference and other fields. We introduce n

Together AI and Pearl Research Labs Team Up to Reduce the Cost of AI Inference

Model ReleasesDGX agent

Together AI and Pearl Research Labs announced a partnership aimed at lowering the cost of AI inference through collaborative research and development efforts. The partnership likely combines Together

Vendor-Conditioned Contrastive Learning for Predicting Organizational Cyber Threat Targets

Model ReleasesDGX agent

arXiv:2012.14425v2 Announce Type: replace-cross Abstract: Cyberattacks cause billions of dollars in damage annually, with malicious hackers often sharing exploit code and techniques on underground for

VerbalValue: A Socially Intelligent Virtual Host for Sales-Driven Live Commerce

Model ReleasesDGX agent

arXiv:2605.14542v1 Announce Type: new Abstract: A skilled live-commerce host is not merely a narrator, but a sales agent who converts viewer curiosity into purchase intent through expert product knowl

What Do AI Agents Talk About? Discourse and Architectural Constraints in the First AI-Only Social Network

Model ReleasesDGX agent

arXiv:2603.07880v5 Announce Type: replace Abstract: Moltbook is the first large-scale social network built for autonomous AI agent-to-agent interaction. Early studies on Moltbook have interpreted its

When Evidence Conflicts: Uncertainty and Order Effects in Retrieval-Augmented Biomedical Question Answering

ResearchDGX agent

arXiv:2605.14115v1 Announce Type: new Abstract: Biomedical retrieval-augmented large language models (LLMs) often face evidence that is incomplete, misleading, or internally contradictory, yet evaluat

Woodelf++: A Fast and Unified Partial Dependence Plot Algorithm for Decision Tree Ensembles

HardwareDGX agent

arXiv:2605.14578v1 Announce Type: new Abstract: Partial Dependence Plots (PDPs) visualize how changes in a single feature affect the average model prediction. They are widely used in practice to inter

14 May 2026

AgenticAITA: A Proof-Of-Concept About Deliberative Multi-Agent Reasoning for Autonomous Trading Systems

SafetyDGX agent

arXiv:2605.12532v1 Announce Type: cross Abstract: Conventional algorithmic trading systems are grounded in deterministic heuristics or offline-trained statistical models that cannot adapt to the seman

Aligning Forest and Trees in Images & Long Captions for Visually Grounded Understanding

SafetyDGX agent

arXiv:2602.02977v2 Announce Type: replace-cross Abstract: Vision-language models such as CLIP often struggle to faithfully understand long, detail-rich captions, relying on dominant scene cues while o

ArcVQ-VAE: A Spherical Vector Quantization Framework with ArcCosine Additive Margin

ResearchDGX agent

arXiv:2605.13517v1 Announce Type: cross Abstract: Vector Quantized Variational Autoencoder (VQ-VAE) has become a fundamental framework for learning discrete representations in image modeling. However,

Asynchronous Reasoning: Training-Free Interactive Thinking LLMs

SafetyDGX agent

arXiv:2512.10931v3 Announce Type: replace Abstract: Many state-of-the-art LLMs are trained to think before giving their answer. Reasoning can greatly improve language model capabilities, but it also m

Attention Once Is All You Need: Efficient Streaming Inference with Stateful Transformers

Model ReleasesDGX agent

arXiv:2605.13784v1 Announce Type: new Abstract: Conventional transformer inference engines are request-driven, paying an O(n) prefill cost on every query. In streaming workloads, where data arrives co

Auditing Sybil: Explaining Deep Lung Cancer Risk Prediction Through Generative Interventional Attributions

SafetyDGX agent

arXiv:2602.02560v2 Announce Type: replace-cross Abstract: Lung cancer remains the leading cause of cancer mortality, driving the development of automated screening tools to alleviate radiologist workl

Backdoor Channels Hidden in Latent Space: Cryptographic Undetectability in Modern Neural Networks

ResearchDGX agent

arXiv:2605.13214v1 Announce Type: cross Abstract: Recent cryptographic results establish that neural networks can be backdoored such that no efficient algorithm can distinguish them from a clean model

Bayesian Nonparametric Mixed-Effect ODEs with Gaussian Processes

ResearchDGX agent

arXiv:2605.13088v1 Announce Type: new Abstract: Dynamical modelling is central to many scientific domains, including pharmacometrics, systems biology, physiology, and epidemiology. In these settings,

Beyond Perplexity: A Geometric and Spectral Study of Low-Rank Pre-Training

ResearchDGX agent

arXiv:2605.13652v1 Announce Type: cross Abstract: Pre-training large language models is dominated by the memory cost of storing full-rank weights, gradients, and optimizer states. Low-rank pre-trainin

Beyond Softmax: A Natural Parameterization for Categorical Random Variables

ResearchDGX agent

arXiv:2509.24728v2 Announce Type: replace Abstract: Latent categorical variables are frequently found in deep learning architectures. They can model actions in discrete reinforcement-learning environm

BrainAnytime: Anatomy-Aware Cross-Modal Pretraining for Brain Image Analysis with Arbitrary Modality Availability

ResearchDGX agent

arXiv:2605.13059v1 Announce Type: new Abstract: Clinical diagnostic workups typically follow a modality escalation pathway: after initial clinical evaluation, clinicians begin with routine structural

Can LLM Agents Simulate Dynamic Networks? A Case Study on Email Networks with Phishing Synthesis

AgentsDGX agent

arXiv:2605.12507v1 Announce Type: cross Abstract: While Large Language Model (LLM) multi-agent systems (MAS) offer a transformative approach to simulating human behavior in complex systems, it remains

Certified Robustness under Heterogeneous Perturbations via Hybrid Randomized Smoothing

SafetyDGX agent

arXiv:2605.12876v1 Announce Type: new Abstract: Randomized smoothing provides strong, model-agnostic robustness certificates, but existing guarantees are limited to single modalities, treating continu

Cloud CISO Perspectives: How Google + Wiz changes multicloud strategy for CISOs

Model ReleasesDGX agent

Welcome to the first Cloud CISO Perspectives for May 2026. Today, Vinod D’Souza, director, Office of the CISO, shares highlights from his RSA Conference fireside chat with Anthony Belfiore, chief stra

ConRetroBert: EMA Stabilized Dual Encoders for Template-Based Single-Step Retrosynthesis

Model ReleasesDGX agent

arXiv:2605.12736v1 Announce Type: new Abstract: Template based single step retrosynthesis predicts reactants by selecting and applying an explicit reaction template, making each prediction traceable t

Constraint-Aware Flow Matching: Decision Aligned End-to-End Training for Constrained Sampling

ApplicationsDGX agent

arXiv:2605.12754v1 Announce Type: new Abstract: Deep generative models provide state-of-the-art performance across a wide array of applications, with recent studies showing increasing applicability fo

Context Training with Active Information Seeking

ResearchDGX agent

arXiv:2605.13050v1 Announce Type: cross Abstract: Most existing large language models (LLMs) are expensive to adapt after deployment, especially when a task requires newly produced information or nich

CoRe-Gen: Robust Spectrum-to-Structure Generation under Imperfect Fingerprint Conditions

Model ReleasesDGX agent

arXiv:2605.12980v1 Announce Type: cross Abstract: Molecular structure elucidation from tandem mass spectra (MS/MS) remains challenging, particularly for de novo generation beyond database coverage. A

CUBic: Coordinated Unified Bimanual Perception and Control Framework

Model ReleasesDGX agent

arXiv:2605.13452v1 Announce Type: cross Abstract: Recent advances in visuomotor policy learning have enabled robots to perform control directly from visual inputs. Yet, extending such end-to-end learn

Deepseek V4 Flash is now free via Nous Portal for a limited time thanks to @novita_labs!

Model ReleasesDGX agent

Nous Research announced that Deepseek V4 Flash is temporarily available for free access through the Nous Portal, courtesy of Nous Research in collaboration with @novita_labs. This limited-time offer p

Digital Twins as Synthetic Controls in Single-Arm Trials

SafetyDGX agent

arXiv:2605.12832v1 Announce Type: cross Abstract: Single-arm trials are an important study design for evaluating drug efficacy and safety without enrolling patients into a control arm. Although they d

DirectTryOn: One-Step Virtual Try-On via Straightened Conditional Transport

ResearchDGX agent

arXiv:2605.12939v1 Announce Type: new Abstract: Recent diffusion- and flow-based VTON methods achieve strong results with pretrained generative models, but their reliance on multi-step sampling incurs

Distribution Shift in Missing Data Imputation: A Risk-Based Perspective and Importance-Weighted Correction under MAR

TutorialsDGX agent

arXiv:2602.06713v2 Announce Type: replace-cross Abstract: Missing data imputation, where a model is trained on observed data to estimate unobserved values, is a fundamental problem in machine learning

Do Heavy Tails Help Diffusion? On the Subtle Trade-off Between Initialization and Training

ApplicationsDGX agent

arXiv:2605.13175v1 Announce Type: new Abstract: Recent works have proposed incorporating heavy-tailed (HT) noise into diffusion- and flow-based generative models, with the goals of better recovering t

Does language matter for spoken word classification? A multilingual generative meta-learning approach

ResearchDGX agent

arXiv:2605.13084v1 Announce Type: cross Abstract: Meta-learning has been shown to have better performance than supervised learning for few-shot monolingual spoken word classification. However, the met

Early Data Exposure Improves Robustness to Subsequent Fine-Tuning

ResearchDGX agent

arXiv:2605.12705v1 Announce Type: new Abstract: How can we train models whose post-trained capabilities survive subsequent fine-tuning? Rather than focusing on downstream interventions to mitigate for

Emergent and Subliminal Misalignment Through the Lens of Data-Mediated Transfer

SafetyDGX agent

arXiv:2605.12798v1 Announce Type: cross Abstract: Fine-tuning LLMs on narrow harmful datasets can induce Emergent Misalignment (EM), where models exhibit misaligned behavior far beyond the fine-tuning

Exact Sequence Interpolation with Transformers

Model ReleasesDGX agent

arXiv:2502.02270v3 Announce Type: replace Abstract: We prove that transformers can exactly interpolate datasets of finite input sequences in R^d, dgeq 2, with corresponding output sequences of smaller

Fireworks Training Platform continues to expand. Today GLM 5.1 LoRA RL is now live via Training API: SFT, DPO, and full RL on a 200K context…

Model ReleasesDGX agent

Fireworks Training Platform continues to expand. Today GLM 5.1 LoRA RL is now live via Training API: SFT, DPO, and full RL on a 200K context window → custom loss functions or smart defaults. No usage

Force-Aware Neural Tangent Kernels for Scalable and Robust Active Learning of MLIPs

Model ReleasesDGX agent

arXiv:2605.13788v1 Announce Type: new Abstract: Active learning for machine-learning interatomic potentials (MLIPs) must address several challenges to be practical: scaling to large candidate pools, l

From Baselines to Transport Geodesics: Axiomatic Attribution via Optimal Generative Flows

ResearchDGX agent

arXiv:2603.05093v2 Announce Type: replace-cross Abstract: Feature attributions often hide a critical modeling choice: they explain a prediction along a counterfactual path from a reference state to an

From Generalist to Specialist Representation

ResearchDGX agent

arXiv:2605.12733v1 Announce Type: cross Abstract: Given a generalist model, learning a task-relevant specialist representation is fundamental for downstream applications. Identifiability, the asymptot

GeoFlowVLM: Geometry-Aware Joint Uncertainty for Frozen Vision-Language Embedding

ResearchDGX agent

arXiv:2605.13352v1 Announce Type: new Abstract: Standard dual-encoder vision-language models that map images and text to deterministic points on a shared unit hypersphere through ell_2 normalization t

Geometric Preconditioning and Curriculum Optimization for Trainable Variational Quantum Regression

Model ReleasesDGX agent

arXiv:2601.11942v3 Announce Type: replace Abstract: Variational quantum circuits are increasingly studied as continuous-function approximators, but quantum regression remains difficult to train when g

GLASS: Global-Local Aggregation for Inference-time Sparsification of LLMs

Local AiDGX agent

arXiv:2508.14302v2 Announce Type: replace-cross Abstract: Inference-time sparsification is a promising path to deploy large language models (LLMs) on resource-constrained devices, yet existing trainin

GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking

TutorialsDGX agent

arXiv:2602.17555v3 Announce Type: replace Abstract: Video reasoning requires a fine-grained understanding of the temporal dependencies and event-level relations between objects and events in videos. C

Hierarchical Transformer Preconditioning for Interactive Physics Simulation

Model ReleasesDGX agent

arXiv:2605.13343v1 Announce Type: cross Abstract: Neural preconditioners for real-time physics simulation offer promising data-driven priors, but they often fail to capture long-range couplings effici

History Anchors: How Prior Behavior Steers LLM Decisions Toward Unsafe Actions

SafetyDGX agent

arXiv:2605.13825v1 Announce Type: new Abstract: Frontier LLMs are increasingly deployed as agents that pick the next action after a long log of prior tool calls produced by the same or a different mod

Implicit Behavioral Decoding from Next-Step Spike Forecasts at Population Scale

Model ReleasesDGX agent

arXiv:2605.12999v1 Announce Type: cross Abstract: Closed-loop brain-computer interfaces often require both a forecast of upcoming neural population activity and a readout of the animal's behavioral st

Imposing Boundary Conditions on Neural Operators via Learned Function Extensions

Model ReleasesDGX agent

arXiv:2602.04923v2 Announce Type: replace Abstract: Neural operators have emerged as powerful surrogates for the solution of partial differential equations (PDEs), yet their ability to handle general,

Improving Classifier-Free Guidance of Flow Matching via Manifold Projection

SafetyDGX agent

arXiv:2601.21892v2 Announce Type: replace-cross Abstract: Classifier-free guidance (CFG) is a widely used technique for controllable generation in diffusion and flow-based models. Despite its empirica

IndicMedDialog: A Parallel Multi-Turn Medical Dialogue Dataset for Accessible Healthcare in Indic Languages

Model ReleasesDGX agent

arXiv:2605.13292v1 Announce Type: cross Abstract: Most existing medical dialogue systems operate in a single-turn question--answering paradigm or rely on template-based datasets, limiting conversation

Information as Maximum-Caliber Deviation: A bridge between Integrated Information Theory and the Free Energy Principle

ResearchDGX agent

arXiv:2605.12536v1 Announce Type: cross Abstract: The Free Energy Principle (FEP) is a leading framework for mathematically modeling self-organization and learning, while Integrated Information Theory

Interesting position paper on agentic AI as a foreseeable pathway to AGI. (bookmark it) There has been strong debate on whether a larger sin…

SafetyDGX agent

Interesting position paper on agentic AI as a foreseeable pathway to AGI. (bookmark it) There has been strong debate on whether a larger single model get us there or a multi-agent system. The authors

Learning Responsibility-Attributed Adversarial Scenarios for Testing Autonomous Vehicles

Model ReleasesDGX agent

arXiv:2605.13751v1 Announce Type: new Abstract: Establishing trustworthy safety assurance for autonomous driving systems (ADSs) requires evidence that failures arise from avoidable system deficiencies

Limitations of Quantum Advantage in Unsupervised Machine Learning

ResearchDGX agent

arXiv:2511.10709v2 Announce Type: replace-cross Abstract: Machine learning models are used for pattern recognition analysis of big data, without direct human intervention. The task of unsupervised lea

← Previous
1…558559560561562…1061
Next →