AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,137 results
Research

When to Think Fast and Slow? AMOR: Adaptive Entropy Gate for Hybrid Models

DGX agent

arXiv:2602.13215v2 Announce Type: replace Abstract: Recurrent-attention hybrids aim to combine the efficiency of recurrence with the expressivity of attention, but existing approaches typically apply

researcharxiv-cs-ai
14 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

You can now power your Hermes Agent, if using OpenAI models, with codex as the runtime for the core tools that it offers, with the flip of a…

DGX agent

Nous Research announced that Hermes Agents powered by OpenAI models can now use Codex as the runtime for executing core tools, enabling improved tool execution capabilities with a simple configuration

agentsnous-research--x
14 May 2026
Safety

A Survey of On-Policy Distillation for Large Language Models

DGX agent

arXiv:2604.00626v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) continue to grow in both capability and cost, transferring frontier capabilities into smaller, deployable stud

safetyarxiv-cs-cl
13 May 2026
Safety

Adaption, co-founded by ex-Cohere VP of AI research Sara Hooker, unveils AutoScientist, which can automate the research loop behind model training and alignment (Russell Brandom/TechCrunch)

DGX agent

Russell Brandom / TechCrunch: Adaption, co-founded by ex-Cohere VP of AI research Sara Hooker, unveils AutoScientist, which can automate the research loop behind model training and alignment — For yea

safetytechmeme
13 May 2026
Industry

China's move to block Meta's Manus acquisition challenges the 'Singapore washing' model, as major Chinese tech companies build significant Singapore presences (Owen Walker/Financial Times)

DGX agent

Owen Walker / Financial Times: China's move to block Meta's Manus acquisition challenges the “Singapore washing” model, as major Chinese tech companies build significant Singapore presences — Beijing'

industrytechmeme
13 May 2026
Safety

Cluster-Aware Neural Collapse Prompt Tuning for Long-Tailed Generalization of Vision-Language Models

DGX agent

arXiv:2605.11939v1 Announce Type: new Abstract: Prompt learning has emerged as an efficient alternative to fine-tuning pre-trained vision-language models (VLMs). Despite its promise, current methods s

safetyarxiv-cs-cv
13 May 2026
Applications

Config, which is building a data layer for robotics foundation models, raised a 27M seed at a 200M+ valuation led by Samsung Venture Investment (Kate Park/TechCrunch)

DGX agent

Kate Park / TechCrunch: Config, which is building a data layer for robotics foundation models, raised a 27M seed at a 200M+ valuation led by Samsung Venture Investment — Asia's push into physical AI i

applicationstechmeme
13 May 2026
Safety

Diffusion-State Policy Optimization for Masked Diffusion Language Models

DGX agent

arXiv:2602.06462v3 Announce Type: replace Abstract: Masked diffusion language models generate text through iterative masked-token filling, but terminal-only rewards on final completions provide coarse

safetyarxiv-cs-cl
13 May 2026
Research

Do Language Models Encode Knowledge of Linguistic Constraint Violations?

DGX agent

arXiv:2605.12055v1 Announce Type: new Abstract: Large Language Models (LLMs) achieve strong linguistic performance, yet their internal mechanisms for producing these predictions remain unclear. We inv

researcharxiv-cs-cl
13 May 2026
Research

DP-{lambda}CGD: Efficient Noise Correlation for Differentially Private Model Training

DGX agent

arXiv:2601.22334v2 Announce Type: replace Abstract: Differentially private stochastic gradient descent (DP-SGD) is the gold standard for training machine learning models with formal differential priva

researcharxiv-cs-lg
13 May 2026
Applications

Efficient LLM-based Advertising via Model Compression and Parallel Verification

DGX agent

arXiv:2605.11582v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable potential in advertising scenarios such as ad creative generation and targeted advertising. However,

applicationsarxiv-cs-cl
13 May 2026
Safety

Enhancing Target-Guided Proactive Dialogue Systems via Conversational Scenario Modeling and Intent-Keyword Bridging

DGX agent

arXiv:2605.11964v1 Announce Type: new Abstract: A target-guided proactive dialogue system aims to steer conversations proactively toward pre-defined targets, such as designated keywords or specific to

safetyarxiv-cs-cl
13 May 2026
Tools

External content is scanned in parallel by ML classifiers and the BrowseSafe model before agents act on it. File connector data is encrypted…

DGX agent

External content is scanned in parallel by ML classifiers and the BrowseSafe model before agents act on it. File connector data is encrypted in transit and at rest, uploaded files automatically delete

toolsperplexity--x
13 May 2026
Applications

From Token to Token Pair: Efficient Prompt Compression for Large Language Models in Clinical Prediction

DGX agent

arXiv:2605.11774v1 Announce Type: new Abstract: By processing electronic health records (EHRs) as natural language sequences, large language models (LLMs) have shown potential in clinical prediction t

applicationsarxiv-cs-cl
13 May 2026
Industry

Google says Chromebooks will get support through their 'existing date commitment', and 'many' models are 'eligible to transition' to the Googlebook experience (Ben Schoon/9to5Google)

DGX agent

Ben Schoon / 9to5Google: Google says Chromebooks will get support through their “existing date commitment”, and “many” models are “eligible to transition” to the Googlebook experience — During The And

industrytechmeme
13 May 2026
Safety

Gradient-Free Noise Optimization for Reward Alignment in Generative Models

DGX agent

arXiv:2605.11347v1 Announce Type: cross Abstract: Existing reward alignment methods for diffusion and flow models rely on multi-step stochastic trajectories, making them difficult to extend to determi

safetyarxiv-cs-cv
13 May 2026
Applications

I don't understand the path forward for Mythos releases. Google & OpenAI will have equivalent models, and they are approaching AI cyber risk…

DGX agent

I don't understand the path forward for Mythos releases. Google & OpenAI will have equivalent models, and they are approaching AI cyber risk guardrails differently, so they will presumably just releas

applicationsethan-mollick--x
13 May 2026
Agents

Joint Learning of Hierarchical Neural Options and Abstract World Model

DGX agent

arXiv:2602.02799v2 Announce Type: replace Abstract: Building agents that can perform new skills by composing existing skills is a long-standing goal of AI agent research. Towards this end, we investig

agentsarxiv-cs-lg
13 May 2026
Safety

Model-based Bootstrap of Controlled Markov Chains

DGX agent

arXiv:2605.12410v1 Announce Type: cross Abstract: We propose and analyze a model-based bootstrap for transition kernels in finite controlled Markov chains (CMCs) with possibly nonstationary or history

safetyarxiv-cs-lg
13 May 2026
Applications

Model-Level GNN Explanations via Rule-to-Graph Readout for Logit Reconstruction

DGX agent

arXiv:2503.09051v2 Announce Type: replace Abstract: We propose a novel model-level GNN explanation framework that shifts the explanation target from class-wise rule extraction to rule-based logit reco

applicationsarxiv-cs-lg
13 May 2026
Industry

New paper: research agenda for secret loyalties Imagine a frontier model that has been trained to covertly advance a specific actor's intere…

DGX agent

New paper: research agenda for secret loyalties Imagine a frontier model that has been trained to covertly advance a specific actor's interests (a nation-state, a CEO, an adversary). @joemkwon argues

industryemad-mostaque--x
13 May 2026
Research

OTT-Vid: Optimal Transport Temporal Token Compression for Video Large Language Models

DGX agent

arXiv:2605.11803v1 Announce Type: new Abstract: As Video Large Language Models (Video-LLMs) scale to longer and more complex videos, their inference cost grows rapidly due to the large volume of visua

researcharxiv-cs-cv
13 May 2026
Applications

PayPal runs 74,000 weekly tasks in Perplexity Enterprise. Teams use it for model validation, channel performance, market trend research, com…

DGX agent

PayPal runs 74,000 weekly tasks in Perplexity Enterprise. Teams use it for model validation, channel performance, market trend research, competitive intelligence, and product analysis. Read the custom

applicationsperplexity--x
13 May 2026
Local Ai

Rank Is Not Capacity: Spectral Occupancy for Latent Graph Models

DGX agent

arXiv:2605.11142v1 Announce Type: new Abstract: Graph representation learning has become a standard approach for analyzing networked data, with latent embeddings widely used for link prediction, commu

local-aiarxiv-cs-lg
13 May 2026
Safety

Red-Teaming Text-to-Image Models via In-Context Experience Replay and Semantic-Preserving Prompt Rewriting

DGX agent

arXiv:2411.16769v3 Announce Type: replace-cross Abstract: Understanding the capabilities of text-to-image (T2I) models in harmful content generation is essential to safety and compliance. However, hum

safetyarxiv-cs-cl
13 May 2026
Safety

Simpson's Paradox in Behavioral Curves: How Aggregation Distorts Parametric Models of User Dynamics

DGX agent

arXiv:2605.11017v1 Announce Type: new Abstract: Behavioral curve modeling -- fitting parametric functions to engagement-versus-exposure data -- is standard practice in recommendation, advertising, and

safetyarxiv-cs-lg
13 May 2026
Safety

The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested

DGX agent

arXiv:2605.11496v1 Announce Type: cross Abstract: Recent published evidence from frontier laboratories shows that contemporary AI models can recognise evaluation contexts, latently represent them, and

safetyarxiv-cs-lg
13 May 2026
Research

TST works in two phases. In phase 1, which covers the first 20-40% of training, the model reads contiguous bags of k tokens, with input embe…

DGX agent

TST works in two phases. In phase 1, which covers the first 20-40% of training, the model reads contiguous bags of k tokens, with input embeddings averaged within each bag, and predicts the next bag o

researchnous-research--x
13 May 2026
Industry

We asked the CEO of HuggingFace @ClementDelangue what the risks of releasing powerful open source models are. He says restricting AI creates…

DGX agent

We asked the CEO of HuggingFace @ClementDelangue what the risks of releasing powerful open source models are. He says restricting AI creates more risk than openness. 'Six, seven years ago, at the time

industryclem-delangue--x
13 May 2026
Tutorials

APCD: Adaptive Path-Contrastive Decoding for Reliable Large Language Model Generation

DGX agent

arXiv:2605.09492v1 Announce Type: cross Abstract: Large language models (LLMs) often suffer from hallucinations due to error accumulation in autoregressive decoding, where suboptimal early token choic

tutorialsarxiv-cs-ai
12 May 2026
Research

ATAAT: Adaptive Threat-Aware Adversarial Tuning Framework against Backdoor Attacks on Vision-Language-Action Models

DGX agent

arXiv:2605.08612v1 Announce Type: new Abstract: Addressing the escalating security vulnerabilities in Vision-Language-Action (VLA) models, this study investigates backdoor attacks targeting the visual

researcharxiv-cs-ro
12 May 2026
Agents

AtteConDA: Attention-Based Conflict Suppression in Multi-Condition Diffusion Models and Synthetic Data Augmentation

DGX agent

arXiv:2605.09425v1 Announce Type: cross Abstract: Recent conditional image generation methods can improve controllability by generating images that are faithful to conditions such as sketches, human p

agentsarxiv-cs-ai
12 May 2026
Applications

Biosignal Fingerprinting: A Cross-Modal PPG-ECG Foundation Model

DGX agent

arXiv:2605.09579v1 Announce Type: cross Abstract: Cardiovascular disease remains the leading cause of global mortality, yet scalable cardiac monitoring is hindered by the gap between diagnostic-rich E

applicationsarxiv-cs-ai
12 May 2026
Research

Can Muon Fine-tune Adam-Pretrained Models?

DGX agent

arXiv:2605.10468v1 Announce Type: new Abstract: Muon has emerged as an efficient alternative to Adam for pretraining, yet remains underused for fine-tuning. A key obstacle is that most open models are

researcharxiv-cs-lg
12 May 2026
Safety

Composing Policy Gradients and Prompt Optimization for Language Model Programs

DGX agent

arXiv:2508.04660v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has proven to be an effective tool for post-training language models (LMs). However, AI systems are increa

safetyarxiv-cs-cl
12 May 2026
Research

Delighted to announce our 3rd world modeling workshop! After NYC and Montreal, we are now headed to Chicago! - August 31st to September 2nd …

DGX agent

Delighted to announce our 3rd world modeling workshop! After NYC and Montreal, we are now headed to Chicago! - August 31st to September 2nd - CfP and details on the website: https://wm-booth.org - @yl

researchyann-lecun--x
12 May 2026
Local Ai

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding

DGX agent

arXiv:2512.06673v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) are rapidly expanding from general video understanding to finer-grained understanding such as spatio-tempor

local-aiarxiv-cs-cv
12 May 2026
Local Ai

DP-LAC: Lightweight Adaptive Clipping for Differentially Private Federated Fine-tuning of Language Models

DGX agent

arXiv:2605.10272v1 Announce Type: cross Abstract: Federated learning (FL) enables the collaborative training of large-scale language models (LLMs) across edge devices while keeping user data on-device

local-aiarxiv-cs-ai
12 May 2026
Safety

EvoStreaming: Your Offline Video Model Is a Natively Streaming Assistant

DGX agent

arXiv:2605.10343v1 Announce Type: cross Abstract: Streaming video understanding demands more than watching longer videos: assistants must decide when to speak in real time, balancing responsiveness ag

safetyarxiv-cs-ai
12 May 2026
Research

FERA: Uncertainty-Aware Federated Reasoning for Large Language Models

DGX agent

arXiv:2605.10082v1 Announce Type: new Abstract: Large language models (LLMs) exhibit strong reasoning capabilities when guided by high-quality demonstrations, yet such data is often distributed across

researcharxiv-cs-cl
12 May 2026
Applications

Forecasting Source Stability in Scientific Experiments using Temporal Learning Models: A Case Study from Tritium Monitoring

DGX agent

arXiv:2605.08140v1 Announce Type: cross Abstract: The Karlsruhe Tritium Neutrino Experiment (KATRIN) aims to measure the absolute neutrino mass with unprecedented sensitivity, requiring precise monito

applicationsarxiv-cs-ai
12 May 2026
Local Ai

Fully Decentralized Cooperative Multi-Agent Reinforcement Learning is A Context Modeling Problem

DGX agent

arXiv:2509.15519v2 Announce Type: replace Abstract: This paper studies fully decentralized cooperative multi-agent reinforcement learning, where each agent solely observes the states, its local action

local-aiarxiv-cs-lg
12 May 2026
Research

Functional Subspace, where language models can use vector algebra to solve problems

DGX agent

arXiv:2602.01687v2 Announce Type: replace-cross Abstract: Large language models (LLMs) were invented for natural language tasks such as translation, but they have proved that they can perform highly c

researcharxiv-cs-ai
12 May 2026
Agents

GenCellAgent: Generalizable, Training-Free Cellular Image Segmentation via Large Language Model Agents

DGX agent

arXiv:2510.13896v2 Announce Type: replace-cross Abstract: Cellular image segmentation is essential for quantitative biology yet remains difficult due to heterogeneous modalities, morphological variabi

agentsarxiv-cs-ai
12 May 2026
Research

Generative Giants, Retrieval Weaklings: Why do Multimodal Large Language Models Fail at Multimodal Retrieval?

DGX agent

arXiv:2512.19115v2 Announce Type: replace Abstract: Despite the remarkable success of multimodal large language models (MLLMs) in generative tasks, we observe that they exhibit a counterintuitive defi

researcharxiv-cs-cv
12 May 2026
Safety

Hierarchical Causal Abduction: A Foundation Framework for Explainable Model Predictive Control

DGX agent

arXiv:2605.10624v1 Announce Type: new Abstract: Model Predictive Control (MPC) is widely used to operate safety-critical infrastructure by predicting future trajectories and optimizing control actions

safetyarxiv-cs-ai
12 May 2026
Tools

Introducing voice finder from Together AI, a new tool to search, filter, and audition 600+ voices across leading TTS models. AI natives can …

DGX agent

Introducing voice finder from Together AI, a new tool to search, filter, and audition 600+ voices across leading TTS models. AI natives can now find the right voice for their app faster by describing

toolstogether-ai--x
12 May 2026
Agents

'It’s pretty easy to change models these days... what creates more lock-in is when state starts to accumulate behind these APIs... memory ha…

DGX agent

'It’s pretty easy to change models these days... what creates more lock-in is when state starts to accumulate behind these APIs... memory has a lot of gravity.' - @hwchase17, Co-Founder & CEO, @LangCh

agentsharrison-chase--x
12 May 2026
← Previous
1…283284285286287…1316
Next →