AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,986 results
Research

A Hormone-inspired Emotion Layer for Transformer language models (HELT)

DGX agent

arXiv:2605.13858v1 Announce Type: cross Abstract: Large Language Models have demonstrated remarkable capabilities in generating contextually relevant and grammatically correct text. However, they fund

researcharxiv-cs-cl
15 May 2026
Research

Causal Foundation Models with Continuous Treatments

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2605.15133v1 Announce Type: new Abstract: Causal inference, estimating causal effects from observational data, is a fundamental tool in many disciplines. Of particular importance across a variet

researcharxiv-cs-lg
15 May 2026
Local Ai

Conditional Attribute Estimation with Autoregressive Sequence Models

DGX agent

arXiv:2605.14004v1 Announce Type: new Abstract: Generative models are often trained with a next-token prediction objective, yet many downstream applications require the ability to estimate or control

local-aiarxiv-cs-ai
15 May 2026
Hardware

EMA: Efficient Model Adaptation for Learning-based Systems

DGX agent

arXiv:2605.13942v1 Announce Type: new Abstract: Machine learning (ML) is increasingly applied to optimize system performance in tasks such as resource management and network simulation. Unlike traditi

hardwarearxiv-cs-lg
15 May 2026
Research

Generalizing Score-based generative models for Heavy-tailed Distributions

DGX agent

arXiv:2603.00772v2 Announce Type: replace-cross Abstract: Score-based generative models (SGMs) have achieved remarkable empirical success, motivating their application to a broad range of data distrib

researcharxiv-cs-lg
15 May 2026
Safety

HeatKV: Head-tuned KV-cache Compression for Visual Autoregressive Modeling

DGX agent

arXiv:2605.14877v1 Announce Type: new Abstract: Visual Autoregressive (VAR) models have recently demonstrated impressive image generation quality while maintaining low latency. However, they suffer fr

safetyarxiv-cs-cv
15 May 2026
Applications

MeMo: Memory as a Model

DGX agent

arXiv:2605.15156v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong performance across a wide range of tasks, but remain frozen after pretraining until subsequent updates. Ma

applicationsarxiv-cs-ai
15 May 2026
Model Releases

Merging Methods for Multilingual Knowledge Editing for Large Language Models: An Empirical Odyssey

DGX agent

arXiv:2605.13919v1 Announce Type: new Abstract: Multilingual knowledge editing (MKE) remains challenging because language-specific edits interfere with one another, even when locate-then-edit methods

model-releasesarxiv-cs-cl
15 May 2026
Safety

Multi-Dimensional Model Integrity and Responsibility Assessment Index and Scoring Framework

DGX agent

arXiv:2605.14550v1 Announce Type: new Abstract: Artificial intelligence in high-stakes tabular domains cannot be evaluated by predictive performance alone, yet current practice still assesses explaina

safetyarxiv-cs-lg
15 May 2026
Local Ai

Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons

DGX agent

arXiv:2603.02115v2 Announce Type: replace-cross Abstract: General-purpose robot reward models are typically trained to predict absolute task progress from expert demonstrations, providing only local,

local-aiarxiv-cs-ai
15 May 2026
Research

TopoPrimer: The Missing Topological Context in Forecasting Models

DGX agent

arXiv:2605.15035v1 Announce Type: new Abstract: We introduce TopoPrimer, a framework that makes the global topological structure of the series population an explicit input to any forecasting model. To

researcharxiv-cs-lg
15 May 2026
Local Ai

Towards Fine-Grained and Verifiable Concept Bottleneck Models

DGX agent

arXiv:2605.14210v1 Announce Type: cross Abstract: Concept Bottleneck Models (CBMs) offer interpretable alternatives to black-box predictors by introducing human-relatable concepts before the final out

local-aiarxiv-cs-ai
15 May 2026
Safety

BEHAVE: A Hybrid AI Framework for Real-Time Modeling of Collective Human Dynamics

DGX agent

arXiv:2605.12730v1 Announce Type: new Abstract: Existing AI systems for modeling human behavior operate at the level of individuals or detect events after they occur. As a result, they systematically

safetyarxiv-cs-ai
14 May 2026
Research

Correct Answers from Sound Reasoning: Verifiable Process Supervision for Language Models

DGX agent

arXiv:2605.12519v1 Announce Type: cross Abstract: Training language models to produce both correct answers and sound reasoning remains an open challenge. Reinforcement learning with verifiable rewards

researcharxiv-cs-ai
14 May 2026
Applications

Decoupled and Divergence-Conditioned Prompt for Multi-domain Dynamic Graph Foundation Models

DGX agent

arXiv:2605.13540v1 Announce Type: cross Abstract: Dynamic graphs are ubiquitous in real-world systems, and building generalizable dynamic Graph Foundation Models has become a frontier in graph learnin

applicationsarxiv-cs-ai
14 May 2026
Agents

deepagents v0.6 is our biggest release yet!!! it’s all about perf - at the model layer w harness profiles, agent layer w code interpreter, a…

DGX agent

deepagents v0.6 is our biggest release yet!!! it’s all about perf - at the model layer w harness profiles, agent layer w code interpreter, and at scale w streaming and delta channels context hub backe

agentsharrison-chase--x
14 May 2026
Applications

Efficient Generative Prediction for EHR Foundation Models: The SCOPE and REACH Estimators

DGX agent

arXiv:2602.03730v2 Announce Type: replace-cross Abstract: Generative foundation models trained on tokenized electronic health record (EHR) timelines show promise for clinical outcome prediction via Mo

applicationsarxiv-cs-lg
14 May 2026
Model Releases

KamonBench: A Grammar-Based Dataset for Evaluating Compositional Factor Recovery in Vision-Language Models

DGX agent

arXiv:2605.13322v1 Announce Type: new Abstract: Kamon (family crests) are an important part of Japanese culture and a natural test case for compositional visual recognition: each crest combines a smal

model-releasesarxiv-cs-cv
14 May 2026
Safety

Large Language Models for Agentic NetOps and AIOps: Architectures, Evaluation, and Safety

DGX agent

arXiv:2605.12729v1 Announce Type: cross Abstract: Large language models are increasingly being used to support network operations (NetOps) and artificial intelligence for IT operations (AIOps), includ

safetyarxiv-cs-ai
14 May 2026
Tutorials

Modeling Heterophily in Multiplex Graphs: An Adaptive Approach for Node Classification

DGX agent

arXiv:2605.12699v1 Announce Type: cross Abstract: Existing multiplex graph models often assume homophily, where connected nodes tend to belong to the same class or share similar attributes. Consequent

tutorialsarxiv-cs-ai
14 May 2026
Research

Plan for Speed: Dilated Scheduling for Masked Diffusion Language Models

DGX agent

arXiv:2506.19037v4 Announce Type: replace-cross Abstract: Masked diffusion language models (MDLMs) promise fast, non-autoregressive text generation, yet existing samplers, which pick tokens to unmask

researcharxiv-cs-ai
14 May 2026
Research

Predict-Project-Renoise: Sampling Diffusion Models under Hard Constraints

DGX agent

arXiv:2601.21033v2 Announce Type: replace Abstract: Diffusion models cannot enforce hard constraints, yet applications in the physical sciences demand exact satisfaction of conservation laws, boundary

researcharxiv-cs-lg
14 May 2026
Safety

Proximal-Based Generative Modeling for Bayesian Inverse Problems

DGX agent

arXiv:2605.13278v1 Announce Type: cross Abstract: Score-based diffusion models demonstrate superior performance in generative tasks but encounter fundamental bottlenecks in inverse problems due to the

safetyarxiv-cs-lg
14 May 2026
Hardware

Together AI STT models now hold the top two spots for transcription speed on the @ArtificialAnlys Speech to Text leaderboard. NVIDIA Parakee…

DGX agent

Together AI STT models now hold the top two spots for transcription speed on the @ArtificialAnlys Speech to Text leaderboard. NVIDIA Parakeet TDT 0.6B V3 on Together AI ranks #1, transcribing 303 seco

hardwaretogether-ai--x
14 May 2026
Model Releases

Topo-R1: Detecting Topological Anomalies via Vision-Language Models

DGX agent

arXiv:2603.13054v2 Announce Type: replace Abstract: Topology is critical in tubular structures such as blood vessels, nerve fibers, and road networks, where connectivity and loop structure govern down

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Training Large Language Models to Predict Clinical Events

DGX agent

arXiv:2605.12817v1 Announce Type: cross Abstract: Longitudinal clinical notes contain rich evidence of how patients evolve over time, but converting this signal into training supervision for clinical

model-releasesarxiv-cs-ai
14 May 2026
Research

Uncovering Symmetry Transfer in Large Language Models via Layer-Peeled Optimization

DGX agent

arXiv:2605.12756v1 Announce Type: cross Abstract: Large language models (LLMs) are pretrained by minimizing the cross-entropy loss for next-token prediction. In this paper, we study whether this optim

researcharxiv-cs-ai
14 May 2026
Local Ai

What’s the best model to use with RAG to create a locally hosted survival and off grid LLm?

DGX agent

This discussion explores which language models work best when combined with RAG (Retrieval-Augmented Generation) for building a locally hosted LLM focused on survival and off-grid living topics. The t

local-air-ollama
14 May 2026
Research

A Mixture Autoregressive Image Generative Model on Quadtree Regions for Gaussian Noise Removal via Variational Bayes and Gradient Methods

DGX agent

arXiv:2605.11585v1 Announce Type: new Abstract: This paper addresses the problem of image denoising for grayscale images. We propose a probabilistic image generative model that combines a quadtree reg

researcharxiv-cs-cv
13 May 2026
Hardware

BOOST: BOttleneck-Optimized Scalable Training Framework for Low-Rank Large Language Models

DGX agent

arXiv:2512.12131v2 Announce Type: replace Abstract: The scale of transformer model pre-training is constrained by the increasing computation and communication cost. Low-rank bottleneck architectures o

hardwarearxiv-cs-lg
13 May 2026
Research

Can Nano Banana 2 Replace Traditional Image Restoration Models? An Evaluation of Its Performance on Image Restoration Tasks

DGX agent

arXiv:2604.03061v2 Announce Type: replace Abstract: Recent advances in generative AI raise the question of whether general-purpose image editing models can serve as unified solutions for image restora

researcharxiv-cs-cv
13 May 2026
Model Releases

CATS: Cascaded Adaptive Tree Speculation for Memory-Limited LLM Inference Acceleration

DGX agent

arXiv:2605.11186v1 Announce Type: new Abstract: Auto-regressive decoding in Large Language Models (LLMs) is inherently memory-bound: every generation step requires loading the model weights and interm

model-releasesarxiv-cs-lg
13 May 2026
Safety

Combining On-Policy Optimization and Distillation for Long-Context Reasoning in Large Language Models

DGX agent

arXiv:2605.12227v1 Announce Type: new Abstract: Adapting large language models (LLMs) to long-context tasks requires post-training methods that remain accurate and coherent over thousands of tokens. E

safetyarxiv-cs-cl
13 May 2026
Research

Concepts in Motion: Temporal Concept Bottleneck Model for Interpretable Video Classification

DGX agent

arXiv:2509.20899v3 Announce Type: replace Abstract: Concept Bottleneck Models (CBMs) enable interpretable image classification by structuring predictions around human-understandable concepts, but exte

researcharxiv-cs-cv
13 May 2026
Model Releases

Correcting Selection Bias in Sparse User Feedback for Large Language Model Quality Estimation: A Multi-Agent Hierarchical Bayesian Approach

DGX agent

arXiv:2605.12177v1 Announce Type: new Abstract: [Abridged] Production LLM deployments receive feedback from a non-random fraction of users: thumbs sit mostly in the tails of the satisfaction distribut

model-releasesarxiv-cs-cl
13 May 2026
Applications

Dynamic Execution Commitment of Vision-Language-Action Models

DGX agent

arXiv:2605.11567v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models predominantly adopt action chunking, i.e., predicting and committing to a short horizon of consecutive low-level act

applicationsarxiv-cs-cv
13 May 2026
Research

From Model Uncertainty to Human Attention: Localization-Aware Visual Cues for Scalable Annotation Review

DGX agent

arXiv:2605.12303v1 Announce Type: cross Abstract: High-quality labeled data is essential for training robust machine learning models, yet obtaining annotations at scale remains expensive. AI-assisted

researcharxiv-cs-cv
13 May 2026
Tools

Introducing AutoScientist. Most model training fails outside of frontier labs. AutoScientist automates the full research loop so it doesn't …

DGX agent

AutoScientist is an AI system developed by Together AI that automates the complete research workflow to enable model training and scientific discovery outside of well-resourced frontier laboratories.

toolstogether-ai--x
13 May 2026
Research

Is Monotonic Sampling Necessary in Diffusion Models?

DGX agent

arXiv:2605.11773v1 Announce Type: new Abstract: Diffusion models generate samples by iteratively denoising a Gaussian prior, traversing a sequence of noise levels that, in every published sampler, dec

researcharxiv-cs-lg
13 May 2026
Research

One-Step Generative Modeling via Wasserstein Gradient Flows

DGX agent

arXiv:2605.11755v1 Announce Type: cross Abstract: Diffusion models and flow-based methods have shown impressive generative capability, especially for images, but their sampling is expensive because it

researcharxiv-cs-cv
13 May 2026
Safety

ORCE: Order-Aware Alignment of Verbalized Confidence in Large Language Models

DGX agent

arXiv:2605.12446v1 Announce Type: cross Abstract: Large language models (LLMs) often produce answers with high certainty even when they are incorrect, making reliable confidence estimation essential f

safetyarxiv-cs-cl
13 May 2026
Research

Overparametrized models with posterior drift

DGX agent

arXiv:2506.23619v2 Announce Type: replace-cross Abstract: This paper investigates the impact of posterior drift on out-of-sample forecasting accuracy in overparametrized machine learning models. We do

researcharxiv-cs-lg
13 May 2026
Agents

Predicting Decisions of AI Agents from Limited Interaction through Text-Tabular Modeling

DGX agent

arXiv:2605.12411v1 Announce Type: cross Abstract: AI agents negotiate and transact in natural language with unfamiliar counterparts: a buyer bot facing an unknown seller, or a procurement assistant ne

agentsarxiv-cs-cl
13 May 2026
Safety

Pretraining Exposure Explains Popularity Judgments in Large Language Models

DGX agent

arXiv:2605.12382v1 Announce Type: new Abstract: Large language models (LLMs) exhibit systematic preferences for well-known entities, a phenomenon often attributed to popularity bias. However, the exte

safetyarxiv-cs-cl
13 May 2026
Safety

Simulation Distillation: Pretraining World Models in Simulation for Rapid Real-World Adaptation

DGX agent

arXiv:2603.15759v2 Announce Type: replace-cross Abstract: Robot learning requires adaptation methods that improve reliably from limited, mixed-quality interaction data. This is especially challenging

safetyarxiv-cs-lg
13 May 2026
Research

Steering Without Breaking: Mechanistically Informed Interventions for Discrete Diffusion Language Models

DGX agent

arXiv:2605.10971v1 Announce Type: cross Abstract: Discrete diffusion language models (DLMs) generate text by iteratively denoising all positions in parallel, offering an alternative to autoregressive

researcharxiv-cs-cl
13 May 2026
Tutorials

What makes a word hard to learn? Modeling L1 influence on English vocabulary difficulty

DGX agent

arXiv:2605.12281v1 Announce Type: new Abstract: What makes a word difficult to learn, and how does the difficulty depend on the learner's native language? We computationally model vocabulary difficult

tutorialsarxiv-cs-cl
13 May 2026
Safety

A Single Neuron Is Sufficient to Bypass Safety Alignment in Large Language Models

DGX agent

arXiv:2605.08513v1 Announce Type: cross Abstract: Safety alignment in language models operates through two mechanistically distinct systems: refusal neurons that gate whether harmful knowledge is expr

safetyarxiv-cs-ai
12 May 2026
← Previous
1…202203204205206…1271
Next →