AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
14 May 2026

When to Think Fast and Slow? AMOR: Adaptive Entropy Gate for Hybrid Models

ResearchDGX agent

arXiv:2602.13215v2 Announce Type: replace Abstract: Recurrent-attention hybrids aim to combine the efficiency of recurrence with the expressivity of attention, but existing approaches typically apply

You can now power your Hermes Agent, if using OpenAI models, with codex as the runtime for the core tools that it offers, with the flip of a…

AgentsDGX agent

Nous Research announced that Hermes Agents powered by OpenAI models can now use Codex as the runtime for executing core tools, enabling improved tool execution capabilities with a simple configuration

13 May 2026

A Survey of On-Policy Distillation for Large Language Models

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety
DGX agent

arXiv:2604.00626v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) continue to grow in both capability and cost, transferring frontier capabilities into smaller, deployable stud

Adaption, co-founded by ex-Cohere VP of AI research Sara Hooker, unveils AutoScientist, which can automate the research loop behind model training and alignment (Russell Brandom/TechCrunch)

SafetyDGX agent

Russell Brandom / TechCrunch: Adaption, co-founded by ex-Cohere VP of AI research Sara Hooker, unveils AutoScientist, which can automate the research loop behind model training and alignment — For yea

China's move to block Meta's Manus acquisition challenges the 'Singapore washing' model, as major Chinese tech companies build significant Singapore presences (Owen Walker/Financial Times)

IndustryDGX agent

Owen Walker / Financial Times: China's move to block Meta's Manus acquisition challenges the “Singapore washing” model, as major Chinese tech companies build significant Singapore presences — Beijing'

Cluster-Aware Neural Collapse Prompt Tuning for Long-Tailed Generalization of Vision-Language Models

SafetyDGX agent

arXiv:2605.11939v1 Announce Type: new Abstract: Prompt learning has emerged as an efficient alternative to fine-tuning pre-trained vision-language models (VLMs). Despite its promise, current methods s

Config, which is building a data layer for robotics foundation models, raised a 27M seed at a 200M+ valuation led by Samsung Venture Investment (Kate Park/TechCrunch)

ApplicationsDGX agent

Kate Park / TechCrunch: Config, which is building a data layer for robotics foundation models, raised a 27M seed at a 200M+ valuation led by Samsung Venture Investment — Asia's push into physical AI i

Diffusion-State Policy Optimization for Masked Diffusion Language Models

SafetyDGX agent

arXiv:2602.06462v3 Announce Type: replace Abstract: Masked diffusion language models generate text through iterative masked-token filling, but terminal-only rewards on final completions provide coarse

Do Language Models Encode Knowledge of Linguistic Constraint Violations?

ResearchDGX agent

arXiv:2605.12055v1 Announce Type: new Abstract: Large Language Models (LLMs) achieve strong linguistic performance, yet their internal mechanisms for producing these predictions remain unclear. We inv

DP-{lambda}CGD: Efficient Noise Correlation for Differentially Private Model Training

ResearchDGX agent

arXiv:2601.22334v2 Announce Type: replace Abstract: Differentially private stochastic gradient descent (DP-SGD) is the gold standard for training machine learning models with formal differential priva

Efficient LLM-based Advertising via Model Compression and Parallel Verification

ApplicationsDGX agent

arXiv:2605.11582v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable potential in advertising scenarios such as ad creative generation and targeted advertising. However,

Enhancing Target-Guided Proactive Dialogue Systems via Conversational Scenario Modeling and Intent-Keyword Bridging

SafetyDGX agent

arXiv:2605.11964v1 Announce Type: new Abstract: A target-guided proactive dialogue system aims to steer conversations proactively toward pre-defined targets, such as designated keywords or specific to

External content is scanned in parallel by ML classifiers and the BrowseSafe model before agents act on it. File connector data is encrypted…

ToolsDGX agent

External content is scanned in parallel by ML classifiers and the BrowseSafe model before agents act on it. File connector data is encrypted in transit and at rest, uploaded files automatically delete

From Token to Token Pair: Efficient Prompt Compression for Large Language Models in Clinical Prediction

ApplicationsDGX agent

arXiv:2605.11774v1 Announce Type: new Abstract: By processing electronic health records (EHRs) as natural language sequences, large language models (LLMs) have shown potential in clinical prediction t

Google says Chromebooks will get support through their 'existing date commitment', and 'many' models are 'eligible to transition' to the Googlebook experience (Ben Schoon/9to5Google)

IndustryDGX agent

Ben Schoon / 9to5Google: Google says Chromebooks will get support through their “existing date commitment”, and “many” models are “eligible to transition” to the Googlebook experience — During The And

Gradient-Free Noise Optimization for Reward Alignment in Generative Models

SafetyDGX agent

arXiv:2605.11347v1 Announce Type: cross Abstract: Existing reward alignment methods for diffusion and flow models rely on multi-step stochastic trajectories, making them difficult to extend to determi

I don't understand the path forward for Mythos releases. Google & OpenAI will have equivalent models, and they are approaching AI cyber risk…

ApplicationsDGX agent

I don't understand the path forward for Mythos releases. Google & OpenAI will have equivalent models, and they are approaching AI cyber risk guardrails differently, so they will presumably just releas

Joint Learning of Hierarchical Neural Options and Abstract World Model

AgentsDGX agent

arXiv:2602.02799v2 Announce Type: replace Abstract: Building agents that can perform new skills by composing existing skills is a long-standing goal of AI agent research. Towards this end, we investig

Model-based Bootstrap of Controlled Markov Chains

SafetyDGX agent

arXiv:2605.12410v1 Announce Type: cross Abstract: We propose and analyze a model-based bootstrap for transition kernels in finite controlled Markov chains (CMCs) with possibly nonstationary or history

Model-Level GNN Explanations via Rule-to-Graph Readout for Logit Reconstruction

ApplicationsDGX agent

arXiv:2503.09051v2 Announce Type: replace Abstract: We propose a novel model-level GNN explanation framework that shifts the explanation target from class-wise rule extraction to rule-based logit reco

New paper: research agenda for secret loyalties Imagine a frontier model that has been trained to covertly advance a specific actor's intere…

IndustryDGX agent

New paper: research agenda for secret loyalties Imagine a frontier model that has been trained to covertly advance a specific actor's interests (a nation-state, a CEO, an adversary). @joemkwon argues

OTT-Vid: Optimal Transport Temporal Token Compression for Video Large Language Models

ResearchDGX agent

arXiv:2605.11803v1 Announce Type: new Abstract: As Video Large Language Models (Video-LLMs) scale to longer and more complex videos, their inference cost grows rapidly due to the large volume of visua

PayPal runs 74,000 weekly tasks in Perplexity Enterprise. Teams use it for model validation, channel performance, market trend research, com…

ApplicationsDGX agent

PayPal runs 74,000 weekly tasks in Perplexity Enterprise. Teams use it for model validation, channel performance, market trend research, competitive intelligence, and product analysis. Read the custom

Rank Is Not Capacity: Spectral Occupancy for Latent Graph Models

Local AiDGX agent

arXiv:2605.11142v1 Announce Type: new Abstract: Graph representation learning has become a standard approach for analyzing networked data, with latent embeddings widely used for link prediction, commu

Red-Teaming Text-to-Image Models via In-Context Experience Replay and Semantic-Preserving Prompt Rewriting

SafetyDGX agent

arXiv:2411.16769v3 Announce Type: replace-cross Abstract: Understanding the capabilities of text-to-image (T2I) models in harmful content generation is essential to safety and compliance. However, hum

Simpson's Paradox in Behavioral Curves: How Aggregation Distorts Parametric Models of User Dynamics

SafetyDGX agent

arXiv:2605.11017v1 Announce Type: new Abstract: Behavioral curve modeling -- fitting parametric functions to engagement-versus-exposure data -- is standard practice in recommendation, advertising, and

The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested

SafetyDGX agent

arXiv:2605.11496v1 Announce Type: cross Abstract: Recent published evidence from frontier laboratories shows that contemporary AI models can recognise evaluation contexts, latently represent them, and

TST works in two phases. In phase 1, which covers the first 20-40% of training, the model reads contiguous bags of k tokens, with input embe…

ResearchDGX agent

TST works in two phases. In phase 1, which covers the first 20-40% of training, the model reads contiguous bags of k tokens, with input embeddings averaged within each bag, and predicts the next bag o

We asked the CEO of HuggingFace @ClementDelangue what the risks of releasing powerful open source models are. He says restricting AI creates…

IndustryDGX agent

We asked the CEO of HuggingFace @ClementDelangue what the risks of releasing powerful open source models are. He says restricting AI creates more risk than openness. 'Six, seven years ago, at the time

12 May 2026

APCD: Adaptive Path-Contrastive Decoding for Reliable Large Language Model Generation

TutorialsDGX agent

arXiv:2605.09492v1 Announce Type: cross Abstract: Large language models (LLMs) often suffer from hallucinations due to error accumulation in autoregressive decoding, where suboptimal early token choic

ATAAT: Adaptive Threat-Aware Adversarial Tuning Framework against Backdoor Attacks on Vision-Language-Action Models

ResearchDGX agent

arXiv:2605.08612v1 Announce Type: new Abstract: Addressing the escalating security vulnerabilities in Vision-Language-Action (VLA) models, this study investigates backdoor attacks targeting the visual

AtteConDA: Attention-Based Conflict Suppression in Multi-Condition Diffusion Models and Synthetic Data Augmentation

AgentsDGX agent

arXiv:2605.09425v1 Announce Type: cross Abstract: Recent conditional image generation methods can improve controllability by generating images that are faithful to conditions such as sketches, human p

Biosignal Fingerprinting: A Cross-Modal PPG-ECG Foundation Model

ApplicationsDGX agent

arXiv:2605.09579v1 Announce Type: cross Abstract: Cardiovascular disease remains the leading cause of global mortality, yet scalable cardiac monitoring is hindered by the gap between diagnostic-rich E

Can Muon Fine-tune Adam-Pretrained Models?

ResearchDGX agent

arXiv:2605.10468v1 Announce Type: new Abstract: Muon has emerged as an efficient alternative to Adam for pretraining, yet remains underused for fine-tuning. A key obstacle is that most open models are

Composing Policy Gradients and Prompt Optimization for Language Model Programs

SafetyDGX agent

arXiv:2508.04660v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has proven to be an effective tool for post-training language models (LMs). However, AI systems are increa

Delighted to announce our 3rd world modeling workshop! After NYC and Montreal, we are now headed to Chicago! - August 31st to September 2nd …

ResearchDGX agent

Delighted to announce our 3rd world modeling workshop! After NYC and Montreal, we are now headed to Chicago! - August 31st to September 2nd - CfP and details on the website: https://wm-booth.org - @yl

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding

Local AiDGX agent

arXiv:2512.06673v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) are rapidly expanding from general video understanding to finer-grained understanding such as spatio-tempor

DP-LAC: Lightweight Adaptive Clipping for Differentially Private Federated Fine-tuning of Language Models

Local AiDGX agent

arXiv:2605.10272v1 Announce Type: cross Abstract: Federated learning (FL) enables the collaborative training of large-scale language models (LLMs) across edge devices while keeping user data on-device

EvoStreaming: Your Offline Video Model Is a Natively Streaming Assistant

SafetyDGX agent

arXiv:2605.10343v1 Announce Type: cross Abstract: Streaming video understanding demands more than watching longer videos: assistants must decide when to speak in real time, balancing responsiveness ag

FERA: Uncertainty-Aware Federated Reasoning for Large Language Models

ResearchDGX agent

arXiv:2605.10082v1 Announce Type: new Abstract: Large language models (LLMs) exhibit strong reasoning capabilities when guided by high-quality demonstrations, yet such data is often distributed across

Forecasting Source Stability in Scientific Experiments using Temporal Learning Models: A Case Study from Tritium Monitoring

ApplicationsDGX agent

arXiv:2605.08140v1 Announce Type: cross Abstract: The Karlsruhe Tritium Neutrino Experiment (KATRIN) aims to measure the absolute neutrino mass with unprecedented sensitivity, requiring precise monito

Fully Decentralized Cooperative Multi-Agent Reinforcement Learning is A Context Modeling Problem

Local AiDGX agent

arXiv:2509.15519v2 Announce Type: replace Abstract: This paper studies fully decentralized cooperative multi-agent reinforcement learning, where each agent solely observes the states, its local action

Functional Subspace, where language models can use vector algebra to solve problems

ResearchDGX agent

arXiv:2602.01687v2 Announce Type: replace-cross Abstract: Large language models (LLMs) were invented for natural language tasks such as translation, but they have proved that they can perform highly c

GenCellAgent: Generalizable, Training-Free Cellular Image Segmentation via Large Language Model Agents

AgentsDGX agent

arXiv:2510.13896v2 Announce Type: replace-cross Abstract: Cellular image segmentation is essential for quantitative biology yet remains difficult due to heterogeneous modalities, morphological variabi

Generative Giants, Retrieval Weaklings: Why do Multimodal Large Language Models Fail at Multimodal Retrieval?

ResearchDGX agent

arXiv:2512.19115v2 Announce Type: replace Abstract: Despite the remarkable success of multimodal large language models (MLLMs) in generative tasks, we observe that they exhibit a counterintuitive defi

Hierarchical Causal Abduction: A Foundation Framework for Explainable Model Predictive Control

SafetyDGX agent

arXiv:2605.10624v1 Announce Type: new Abstract: Model Predictive Control (MPC) is widely used to operate safety-critical infrastructure by predicting future trajectories and optimizing control actions

Introducing voice finder from Together AI, a new tool to search, filter, and audition 600+ voices across leading TTS models. AI natives can …

ToolsDGX agent

Introducing voice finder from Together AI, a new tool to search, filter, and audition 600+ voices across leading TTS models. AI natives can now find the right voice for their app faster by describing

'It’s pretty easy to change models these days... what creates more lock-in is when state starts to accumulate behind these APIs... memory ha…

AgentsDGX agent

'It’s pretty easy to change models these days... what creates more lock-in is when state starts to accumulate behind these APIs... memory has a lot of gravity.' - @hwchase17, Co-Founder & CEO, @LangCh

Lakestream: A Consistent and Brokerless Data Plane for Large Foundation Model Training

ResearchDGX agent

arXiv:2605.09994v1 Announce Type: cross Abstract: Modern Large Foundation Model (LFM) training has transformed the data pipeline from a static ingestion layer into a dynamic component that must co-evo

LAQuant: A Simple Overhead-free Large Reasoning Model Quantization by Layer-wise Lookahead Loss

SafetyDGX agent

arXiv:2605.08755v1 Announce Type: new Abstract: Large reasoning models (LRMs) reach competition-level math and coding accuracy via long autoregressive decoding, making per-token decoding cost a primar

Large Language Models over Networks: Collaborative Intelligence under Resource Constraints

Local AiDGX agent

arXiv:2605.08626v1 Announce Type: cross Abstract: Large language models (LLMs) are transforming society, powering applications from smartphone assistants to autonomous driving. Yet cloud-based LLM ser

Learning predictive models for combinations of heterogeneous proteomic data sources

ResearchDGX agent

arXiv:2605.08958v1 Announce Type: new Abstract: Multiple technologies that measure expression levels of protein mixtures in the human body offer a potential for detection and understanding the disease

LoopVLA: Learning Sufficiency in Recurrent Refinement for Vision-Language-Action Models

SafetyDGX agent

arXiv:2605.09948v1 Announce Type: new Abstract: Current Vision-Language-Action (VLA) models typically treat the deepest representation of a vision-language backbone as universally optimal for action p

Majority Bit-Aware Watermarking For Large Language Models

ResearchDGX agent

arXiv:2508.03829v2 Announce Type: replace Abstract: The growing deployment of Large Language Models (LLMs) has raised concerns about their misuse in generating harmful or deceptive content. To address

mHC-SSM: Manifold-Constrained Hyper-Connections for State Space Language Models with Stream-Specialized Adapters

HardwareDGX agent

arXiv:2605.08300v1 Announce Type: cross Abstract: Manifold-Constrained Hyper-Connections (mHC) introduce a stability-motivated variant of multi stream residual mixing by constraining residual stream m

Mitigating Data Scarcity in Spaceflight Applications for Offline Reinforcement Learning Using Physics-Informed Deep Generative Models

SafetyDGX agent

arXiv:2604.02438v2 Announce Type: replace Abstract: The deployment of reinforcement learning (RL)-based controllers on physical systems is often limited by poor generalization to real-world scenarios,

Model-Reference Adaptive Flight Control of the 95-mg Bee++

ResearchDGX agent

arXiv:2605.08525v1 Announce Type: new Abstract: We introduce a model-reference adaptive control (MRAC) architecture for high-performance positional tracking of the Bee++, a 95-mg insect-scale flapping

MolWorld: Molecule World Models for Actionable Molecular Optimization

Local AiDGX agent

arXiv:2605.08954v1 Announce Type: cross Abstract: Molecular optimization in drug discovery aims to discover molecules with improved target properties, but practical lead optimization often requires mo

Monocular Biomechanical Tracking of Fingers with Inverse Kinematics to Foundation Models

SafetyDGX agent

arXiv:2605.09258v1 Announce Type: cross Abstract: Accurate hand and finger tracking from video has significant clinical applications for monitoring activities of daily living and measuring range of mo

Not-So-Strange Love: Language Models and Generative Linguistic Theories are More Compatible than They Appear

ResearchDGX agent

arXiv:2605.10061v1 Announce Type: cross Abstract: Futrell and Mahowald (2025) frame the success of neural language models (LMs) as supporting gradient, usage-based linguistic theories. I argue that LM

← Previous
1…218219220221222…1017
Next →