AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,046 results
Applications

Graph Concept Bottleneck Models

DGX agent

arXiv:2508.14255v2 Announce Type: replace Abstract: Concept Bottleneck Models (CBMs) provide explicit interpretations for deep neural networks through concepts and allow intervention with concepts to

applicationsarxiv-cs-lg
4 May 2026
Local Ai

i thought on device ai was stupid but now my local voice model turns 4x faster on my corporate m4 max laptop than my m2 max personal laptop …

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

i thought on device ai was stupid but now my local voice model turns 4x faster on my corporate m4 max laptop than my m2 max personal laptop now i want to have a beefier computer to run the transcripti

local-aiclem-delangue--x
4 May 2026
Safety

Impact of Task Phrasing on Presumptions in Large Language Models

DGX agent

arXiv:2605.00436v1 Announce Type: new Abstract: Concerns with the safety and reliability of applying large-language models (LLMs) in unpredictable real-world applications motivate this study, which ex

safetyarxiv-cs-cl
4 May 2026
Safety

PrefMoE: Robust Preference Modeling with Mixture-of-Experts Reward Learning

DGX agent

arXiv:2605.00384v1 Announce Type: new Abstract: Preference-based reinforcement learning offers a scalable alternative to manual reward engineering by learning reward structures from comparative feedba

safetyarxiv-cs-ro
4 May 2026
Model Releases

Putting HUMANS first: Efficient LAM Evaluation with Human Preference Alignment

DGX agent

arXiv:2605.00022v1 Announce Type: new Abstract: The rapid proliferation of large audio models (LAMs) demands efficient approaches for model comparison, yet comprehensive benchmarks are costly. To fill

model-releasesarxiv-cs-cl
4 May 2026
Research

Representation in large language models

DGX agent

arXiv:2501.00885v2 Announce Type: replace Abstract: The extraordinary success of recent Large Language Models (LLMs) on a diverse array of tasks has led to an explosion of scientific and philosophical

researcharxiv-cs-cl
4 May 2026
Safety

Resting Neurons, Active Insights: Robustify Activation Sparsity for Large Language Models

DGX agent

arXiv:2512.12744v3 Announce Type: replace Abstract: Activation sparsity offers a compelling route to accelerate large language model (LLM) inference by selectively suppressing hidden activations, yet

safetyarxiv-cs-lg
4 May 2026
Safety

Unlearning What Matters: Token-Level Attribution for Precise Language Model Unlearning

DGX agent

arXiv:2605.00364v1 Announce Type: new Abstract: Machine unlearning has emerged as a critical capability for addressing privacy, safety, and regulatory concerns in large language models (LLMs). Existin

safetyarxiv-cs-cl
4 May 2026
Local Ai

Built an open-source cognitive OS — persistent memory, 24/7 runtime, bring your own model

DGX agent

An open-source locally-run conversational AI that moves beyond simple request-response models by implementing persistent memory, belief, and self-reflection. The system stores all user profiles, memor

local-air-ollama
3 May 2026
Agents

@SakanaAILabs Sakana Fugu: A Multi-Agent Orchestration System as a Foundation Model https://sakana.ai/fugu-beta/

DGX agent

Sakana AI Labs has developed Fugu, a multi-agent orchestration system designed to function as a foundation model. The system likely enables coordinated interaction between multiple AI agents to improv

agentsdavid-ha--x
3 May 2026
Industry

i keep thinking i want the models to be cheaper/faster more than i want them to be smarter but it seems that just being smarter is still the…

DGX agent

Sam Altman reflects on the trade-off between model performance improvements and practical considerations like cost and speed, suggesting that increased intelligence remains a primary focus despite use

industrysam-altman--x
2 May 2026
Local Ai

RTX 5080 with 16 GB VRAM, 64 GB RAM best quantized model for programming?

DGX agent

For programming tasks with an RTX 5080 (16GB VRAM) and 64GB RAM, optimal quantized models include Qwen 3 14B at Q6 quantization, Llama 3.1 13B at Q8, or DeepSeek R1 Distill 14B Q4, all of which fit co

local-air-ollama
2 May 2026
Agents

A Grid-Aware Agent-Based Model for Analyzing Electric Vehicle Charging Systems

DGX agent

arXiv:2604.27849v1 Announce Type: new Abstract: This paper presents a configurable, grid-aware Agent-Based Model (ABM) for the systematic analysis of electric vehicle (EV) charging systems under confi

agentsarxiv-cs-ai
1 May 2026
Safety

AI Models for Depressive Disorder Detection and Diagnosis: A Review

DGX agent

arXiv:2508.12022v2 Announce Type: replace Abstract: Major Depressive Disorder is one of the leading causes of disability worldwide, yet its diagnosis still depends largely on subjective clinical asses

safetyarxiv-cs-ai
1 May 2026
Agents

Backdoor Attacks on Prompt-Driven Video Segmentation Foundation Models

DGX agent

arXiv:2512.22046v2 Announce Type: replace Abstract: Prompt-driven Video Segmentation Foundation Models (VSFMs), such as SAM2, are increasingly used in applications including autonomous driving and dig

agentsarxiv-cs-cv
1 May 2026
Model Releases

Beyond the Mean: Within-Model Reliable Change Detection for LLM Evaluation

DGX agent

arXiv:2604.27405v1 Announce Type: cross Abstract: We adapted the Reliable Change Index (RCI; Jacobson and Truax, 1991) from clinical psychology to item-level LLM version comparison on 2,000 MMLU-Pro i

model-releasesarxiv-cs-ai
1 May 2026
Research

Budget-Constrained Online Retrieval-Augmented Generation: The Chunk-as-a-Service Model

DGX agent

arXiv:2604.26981v1 Announce Type: cross Abstract: Large Language Models (LLMs) have revolutionized the field of natural language processing. However, they exhibit some limitations, including a lack of

researcharxiv-cs-lg
1 May 2026
Local Ai

CasLayout: Cascaded 3D Layout Diffusion for Indoor Scene Synthesis with Implicit Relation Modeling

DGX agent

arXiv:2604.27361v1 Announce Type: new Abstract: Synthesizing realistic 3D indoor scenes remains challenging due to data scarcity and the difficulty of simultaneously enforcing global architectural con

local-aiarxiv-cs-cv
1 May 2026
Model Releases

Cross-Lingual Response Consistency in Large Language Models: An ILR-Informed Evaluation of Claude Across Six Languages

DGX agent

arXiv:2604.27137v1 Announce Type: new Abstract: This paper introduces a systematic evaluation framework grounded in the Interagency Language Roundtable (ILR) Skill Level Descriptions and applies it to

model-releasesarxiv-cs-cl
1 May 2026
Local Ai

Detecting is Easy, Adapting is Hard: Local Expert Growth for Visual Model-Based Reinforcement Learning under Distribution Shift

DGX agent

arXiv:2604.27411v1 Announce Type: new Abstract: Visual model-based reinforcement learning (MBRL) agents can perform well on the training distribution, but often break down once the test environment sh

local-aiarxiv-cs-lg
1 May 2026
Tutorials

Efficient Sparse Selective-Update RNNs for Long-Range Sequence Modeling

DGX agent

arXiv:2603.02226v2 Announce Type: replace Abstract: Real-world sequential signals, such as audio or video, contain critical information that is often embedded within long periods of silence or noise.

tutorialsarxiv-cs-lg
1 May 2026
Safety

Exploring Applications of Transfer-State Large Language Models: Cognitive Profiling and Socratic AI Tutoring

DGX agent

arXiv:2604.27454v1 Announce Type: new Abstract: Large language models (LLMs) sometimes exhibit qualitative shifts in response style under sustained self-referential dialogue conditions (Berg et al., 2

safetyarxiv-cs-cl
1 May 2026
Research

Learning to Spend: Model Predictive Control for Budgeting under Non-Stationary Returns

DGX agent

arXiv:2604.27186v1 Announce Type: cross Abstract: We study finite-horizon budget allocation as a closed-loop economic control problem and evaluate receding-horizon Model Predictive Control (MPC) relat

researcharxiv-cs-ai
1 May 2026
Model Releases

Predicting Upcoming Stuttering Events from Three-Second Audio: Stratified Evaluation Reveals Severity-Selective Precursors, and the Model Deploys Fully On-Device

DGX agent

arXiv:2604.27279v1 Announce Type: cross Abstract: Audio-based stuttering systems to date have been trained for detection -- what disfluency is present now -- leaving prediction, the capability needed

model-releasesarxiv-cs-lg
1 May 2026
Research

Proactive Dialogue Model with Intent Prediction

DGX agent

arXiv:2604.27379v1 Announce Type: new Abstract: Dialogue models are inherently reactive, responding to the current user turn without anticipating upcoming intents, which leads to redundant interaction

researcharxiv-cs-cl
1 May 2026
Research

Repetition over Diversity: High-Signal Data Filtering for Sample-Efficient German Language Modeling

DGX agent

arXiv:2604.28075v1 Announce Type: cross Abstract: Recent research has shown that filtering massive English web corpora into high-quality subsets significantly improves training efficiency. However, fo

researcharxiv-cs-ai
1 May 2026
Research

Sampler-Robust Optimization under Generative Models

DGX agent

arXiv:2604.27447v1 Announce Type: cross Abstract: Modern stochastic optimization pipelines increasingly rely on learned generative models to represent uncertainty, while downstream decisions are evalu

researcharxiv-cs-ai
1 May 2026
Research

Semantic Structure of Feature Space in Large Language Models

DGX agent

arXiv:2604.27169v1 Announce Type: new Abstract: We show that the geometric relations between semantic features in large language models' hidden states closely mirror human psychological associations.

researcharxiv-cs-cl
1 May 2026
Research

Simulating clinical interventions with a generative multimodal model of human physiology

DGX agent

arXiv:2604.27899v1 Announce Type: new Abstract: Understanding how human health changes over time, and why responses to interventions vary between individuals, remains a central challenge in medicine.

researcharxiv-cs-ai
1 May 2026
Model Releases

The TEA Nets framework combines AI and cognitive network science to model targets, events and actors in text

DGX agent

arXiv:2604.27673v1 Announce Type: new Abstract: We introduce Target-Event-Agent Networks (TEA Nets) as a computational framework to extract subjects (``Agents'), verbs (``Events'), and objects (``Targ

model-releasesarxiv-cs-ai
1 May 2026
Safety

Vision-Language Models Mistake Head Orientation for Gaze Direction: Nonverbal Conversation Cues

DGX agent

arXiv:2506.05412v3 Announce Type: replace-cross Abstract: Where someone looks is a nonverbal communication cue that children and adults readily use. How well can Vision-Language Models (VLMs) infer ga

safetyarxiv-cs-cl
1 May 2026
Tutorials

You don't have to choose between either. It's best to use a combination of them. My advice is to learn how to use a few of these models in d…

DGX agent

You don't have to choose between either. It's best to use a combination of them. My advice is to learn how to use a few of these models in different harnesses. Learn to combine their strengths. Open-w

tutorialsdair-ai--x
1 May 2026
Research

A Dual-Task Paradigm to Investigate Sentence Comprehension Strategies in Language Models

DGX agent

arXiv:2604.26351v1 Announce Type: new Abstract: Language models (LMs) behave more like humans when their cognitive resources are restricted, particularly in predicting sentence processing costs such a

researcharxiv-cs-cl
30 Apr 2026
Research

ComboStoc: Combinatorial Stochasticity for Diffusion Generative Models

DGX agent

arXiv:2405.13729v3 Announce Type: replace-cross Abstract: In this paper, we study an under-explored but important factor of diffusion generative models, i.e., the combinatorial complexity. Data sample

researcharxiv-cs-ai
30 Apr 2026
Safety

Delta Score Matters! Spatial Adaptive Multi Guidance in Diffusion Models

DGX agent

arXiv:2604.26503v1 Announce Type: new Abstract: Diffusion models have achieved remarkable success in synthesizing complex static and temporal visuals, a breakthrough largely driven by Classifier-Free

safetyarxiv-cs-cv
30 Apr 2026
Safety

DORA: A Scalable Asynchronous Reinforcement Learning System for Language Model Training

DGX agent

arXiv:2604.26256v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a critical paradigm for LLM post-training, yet the rollout phase -- accounting for 50--80% of total step time --

safetyarxiv-cs-lg
30 Apr 2026
Safety

From Prompt Risk to Response Risk: Paired Analysis of Safety Behavior of Large Language Model

DGX agent

arXiv:2604.26052v1 Announce Type: new Abstract: Safety evaluations of large language models (LLMs) typically report binary outcomes such as attack success rate, refusal rate, or harmful/not-harmful re

safetyarxiv-cs-cl
30 Apr 2026
Research

HIVE: Hidden-Evidence Verification for Hallucination Detection in Diffusion Large Language Models

DGX agent

arXiv:2604.26139v1 Announce Type: new Abstract: Diffusion large language models generate text through multi-step denoising, where hallucination signals may emerge throughout the trajectory rather than

researcharxiv-cs-cl
30 Apr 2026
Tools

i havent done the work to compare it to peers but i'm just excited that we have a base model and honestly for all the people that complained…

DGX agent

i havent done the work to compare it to peers but i'm just excited that we have a base model and honestly for all the people that complained about the death of the completions API (@deepfates ? or dee

toolsswyx--x
30 Apr 2026
Industry

it is quite significant that Musk admitted on the stand that xAI is distilling OpenAI models to train xAI, and that it is using OpenAI's tec…

DGX agent

Elon Musk allegedly admitted during legal testimony that xAI is distilling OpenAI models to train its own systems and utilizing OpenAI's technology, according to a post by Clem Delangue discussing sta

industryclem-delangue--x
30 Apr 2026
Model Releases

LLM-Flax : Generalizable Robotic Task Planning via Neuro-Symbolic Approaches with Large Language Models

DGX agent

arXiv:2604.26569v1 Announce Type: new Abstract: Deploying a neuro-symbolic task planner on a new domain today requires significant manual effort: a domain expert must author relaxation and complementa

model-releasesarxiv-cs-ro
30 Apr 2026
Agents

Our agent harness makes models inside Cursor faster, smarter, and more token-efficient. Here's how we test improvements to the harness, moni…

DGX agent

Our agent harness makes models inside Cursor faster, smarter, and more token-efficient. Here's how we test improvements to the harness, monitor and repair degradations, and customize it for different

agentscursor--x
30 Apr 2026
Research

STARFlow-V: End-to-End Video Generative Modeling with Normalizing Flows

DGX agent

Normalizing flows (NFs) are end-to-end likelihood-based generative models for continuous data, and have recently regained attention with encouraging progress on image generation. Yet in the video gene

researchapple-ml-research
30 Apr 2026
Research

Stochastic Scaling Limits and Synchronization by Noise in Deep Transformer Models

DGX agent

arXiv:2604.26898v1 Announce Type: cross Abstract: We prove pathwise convergence of the layerwise evolution of tokens in a finite-depth, finite-width transformer model with MultiLayer Perceptron (MLP)

researcharxiv-cs-lg
30 Apr 2026
Research

Text-Utilization for Encoder-dominated Speech Recognition Models

DGX agent

arXiv:2604.26514v1 Announce Type: cross Abstract: This paper investigates efficient methods for utilizing text-only data to improve speech recognition, focusing on encoder-dominated models that facili

researcharxiv-cs-ai
30 Apr 2026
Research

Understanding DNNs in Feature Interaction Models: A Dimensional Collapse Perspective

DGX agent

arXiv:2604.26489v1 Announce Type: new Abstract: DNNs have gained widespread adoption in feature interaction recommendation models. However, there has been a longstanding debate on their roles. On one

researcharxiv-cs-lg
30 Apr 2026
Model Releases

BifDet: A 3D Bifurcation Detection Dataset for Airway-Tree Modeling

DGX agent

arXiv:2604.24999v1 Announce Type: new Abstract: Thoracic Computed Tomography (CT) scans offer detailed insights into the intricate branching network of the airway tree, which is essential for understa

model-releasesarxiv-cs-cv
29 Apr 2026
Agents

for a harness to be effective, it needs to be custom fit to a use case. we're offering harness variants custom fit to models!

DGX agent

LangChain is offering multiple harness variants that are custom-fitted to specific use cases and models, recognizing that effective harness implementation requires tailored configuration rather than o

agentsharrison-chase--x
29 Apr 2026
← Previous
1…239240241242243…1272
Next →