AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,597 results
5 May 2026

LUMINA: A Grid Foundation Model for Benchmarking AC Optimal Power Flow Surrogate Learning

Model ReleasesDGX agent

arXiv:2605.02133v1 Announce Type: new Abstract: AC optimal power flow (ACOPF) is foundational yet computationally expensive in power grid operations, driving learning-based surrogates for large-scale

Medmarks: A Comprehensive Open-Source LLM Benchmark Suite for Medical Tasks

Model ReleasesDGX agent

arXiv:2605.01417v1 Announce Type: new Abstract: Evaluating large language models (LLMs) for medical applications remains challenging due to benchmark saturation, limited data accessibility, and insuff

Nace.AI, which lets companies build specialized AI models tailored to their business' mission and language, raised a $21.5M seed led by Walden Catalyst (Chris Metinko/Axios)

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications
DGX agent

Chris Metinko / Axios: Nace.AI, which lets companies build specialized AI models tailored to their business' mission and language, raised a 21.5M seed led by Walden Catalyst — Nace.AI, which lets busi

On Training Large Language Models for Long-Horizon Tasks: An Empirical Study of Horizon Length

ResearchDGX agent

arXiv:2605.02572v1 Announce Type: cross Abstract: Large language models (LLMs) have shown promise as interactive agents that solve tasks through extended sequences of environment interactions. While p

RADMI: Latent Information Aggregation as a Proxy for Model Uncertainty

Model ReleasesDGX agent

arXiv:2605.01502v1 Announce Type: new Abstract: Epistemic uncertainty estimation is essential for identifying regions where deep learning system outputs may be unreliable. However, existing approaches

RefusalGuard: Geometry-Preserving Fine-Tuning for Safety in LLMs

Model ReleasesDGX agent

arXiv:2605.01913v1 Announce Type: cross Abstract: Fine-tuning safety-aligned language models for downstream tasks often leads to substantial degradation of refusal behavior, making models vulnerable t

Silicon Showdown: Performance, Efficiency, and Ecosystem Barriers in Consumer-Grade LLM Inference

Model ReleasesDGX agent

arXiv:2605.00519v2 Announce Type: cross Abstract: The operational landscape of local Large Language Model (LLM) inference has shifted from lightweight models to datacenter-class weights exceeding 70B

Task-Related Token Compression in Multimodal Large Language Models from an Explainability Perspective

SafetyDGX agent

arXiv:2506.01097v2 Announce Type: replace Abstract: Existing Multimodal Large Language Models (MLLMs) process a large number of visual tokens, leading to significant computational costs and inefficien

TCDA: Thread-Constrained Discourse-Aware Modeling for Conversational Sentiment Quadruple Analysis

Model ReleasesDGX agent

arXiv:2605.01717v1 Announce Type: new Abstract: Conversational Aspect-based Sentiment Quadruple Analysis (DiaASQ) needs to capture the complex interrelationships in multiple rounds of dialogues. Exist

TetraJet-v2: Accurate NVFP4 Training for Large Language Models with Oscillation Suppression and Outlier Control

ResearchDGX agent

arXiv:2510.27527v2 Announce Type: replace Abstract: Large Language Models (LLMs) training is prohibitively expensive, driving interest in low-precision fully-quantized training (FQT). While novel 4-bi

The Cylindrical Representation Hypothesis for Language Model Steering

ResearchDGX agent

arXiv:2605.01844v1 Announce Type: new Abstract: Steering is a widely used technique for controlling large language models, yet its effects are often unstable and hard to predict. Existing theoretical

TOC-SR: Task-Optimal Compact diffusion for Image Super Resolution

Model ReleasesDGX agent

arXiv:2605.02767v1 Announce Type: new Abstract: Diffusion models have recently demonstrated strong performance for image restoration tasks, including super-resolution. However, their large model size

Turning Drift into Constraint: Robust Reasoning Alignment in Non-Stationary Environments

Model ReleasesDGX agent

arXiv:2510.04142v2 Announce Type: replace Abstract: This paper identifies a critical yet underexplored challenge in reasoning alignment from multiple multi-modal large language models (MLLMs): In non-

Validation of an AI-based end-to-end model for prostate pathology using long-term archived routine samples

ResearchDGX agent

arXiv:2605.02614v1 Announce Type: new Abstract: Artificial intelligence (AI) is becoming a clinical tool for prostate pathology, but generalization across variations in sample preparation and preserva

Where Do Prompt Perturbations Break Generation? A Segment-Level View of Robustness in LoRA-Tuned Language Models

ResearchDGX agent

arXiv:2605.01605v1 Announce Type: new Abstract: Large language models are sensitive to minor prompt perturbations, yet existing robustness methods usually enforce consistency at the whole-sequence lev

Zero-Shot Adaptation of Behavioral Foundation Models to Unseen Dynamics

SafetyDGX agent

arXiv:2505.13150v2 Announce Type: replace Abstract: Behavioral Foundation Models (BFMs) proved successful in producing policies for arbitrary tasks in a zero-shot manner, requiring no test-time traini

4 May 2026

A unified perspective on fine-tuning and sampling with diffusion and flow models

SafetyDGX agent

arXiv:2605.00229v1 Announce Type: cross Abstract: We study the problem of training diffusion and flow generative models to sample from target distributions defined by an exponential tilting of a base

Agentic development demands a multi-model strategy — and the governance to match

AgentsDGX agent

The rapid rise of agentic development has radically transformed the software engineering landscape, compelling enterprises to embrace a multi-model AI ecosystem. The pace of change in software develop

Better Models Won’t Save Your Agent

AgentsDGX agent

This article likely argues that improving the underlying language models is insufficient for building effective AI agents, and that other critical components—such as knowledge retrieval, context manag

Comparative Analysis of Polygon-Based and Global Machine Learning Models for Bus Occupancy Prediction

Local AiDGX agent

arXiv:2605.00083v1 Announce Type: new Abstract: Accurate forecasting of bus ridership (passengers numbers) is crucial for efficient management and optimization of public transport systems. Traditional

Diversity in Large Language Models under Supervised Fine-Tuning

ResearchDGX agent

arXiv:2605.00195v1 Announce Type: new Abstract: Supervised Fine-Tuning (SFT) is essential for aligning Large Language Models (LLMs) with user intent, yet it is believed to suppress generative diversit

Foundation AI Models for Aerosol Optical Depth Estimation from PACE Satellite Data

SafetyDGX agent

arXiv:2605.00678v1 Announce Type: new Abstract: Aerosol Optical Depth (AOD) retrieval is essential for Earth observation, supporting applications from air quality monitoring to climate studies. Conven

Graph Concept Bottleneck Models

ApplicationsDGX agent

arXiv:2508.14255v2 Announce Type: replace Abstract: Concept Bottleneck Models (CBMs) provide explicit interpretations for deep neural networks through concepts and allow intervention with concepts to

i thought on device ai was stupid but now my local voice model turns 4x faster on my corporate m4 max laptop than my m2 max personal laptop …

Local AiDGX agent

i thought on device ai was stupid but now my local voice model turns 4x faster on my corporate m4 max laptop than my m2 max personal laptop now i want to have a beefier computer to run the transcripti

Impact of Task Phrasing on Presumptions in Large Language Models

SafetyDGX agent

arXiv:2605.00436v1 Announce Type: new Abstract: Concerns with the safety and reliability of applying large-language models (LLMs) in unpredictable real-world applications motivate this study, which ex

PrefMoE: Robust Preference Modeling with Mixture-of-Experts Reward Learning

SafetyDGX agent

arXiv:2605.00384v1 Announce Type: new Abstract: Preference-based reinforcement learning offers a scalable alternative to manual reward engineering by learning reward structures from comparative feedba

Putting HUMANS first: Efficient LAM Evaluation with Human Preference Alignment

Model ReleasesDGX agent

arXiv:2605.00022v1 Announce Type: new Abstract: The rapid proliferation of large audio models (LAMs) demands efficient approaches for model comparison, yet comprehensive benchmarks are costly. To fill

Representation in large language models

ResearchDGX agent

arXiv:2501.00885v2 Announce Type: replace Abstract: The extraordinary success of recent Large Language Models (LLMs) on a diverse array of tasks has led to an explosion of scientific and philosophical

Resting Neurons, Active Insights: Robustify Activation Sparsity for Large Language Models

SafetyDGX agent

arXiv:2512.12744v3 Announce Type: replace Abstract: Activation sparsity offers a compelling route to accelerate large language model (LLM) inference by selectively suppressing hidden activations, yet

Unlearning What Matters: Token-Level Attribution for Precise Language Model Unlearning

SafetyDGX agent

arXiv:2605.00364v1 Announce Type: new Abstract: Machine unlearning has emerged as a critical capability for addressing privacy, safety, and regulatory concerns in large language models (LLMs). Existin

3 May 2026

Built an open-source cognitive OS — persistent memory, 24/7 runtime, bring your own model

Local AiDGX agent

An open-source locally-run conversational AI that moves beyond simple request-response models by implementing persistent memory, belief, and self-reflection. The system stores all user profiles, memor

@SakanaAILabs Sakana Fugu: A Multi-Agent Orchestration System as a Foundation Model https://sakana.ai/fugu-beta/

AgentsDGX agent

Sakana AI Labs has developed Fugu, a multi-agent orchestration system designed to function as a foundation model. The system likely enables coordinated interaction between multiple AI agents to improv

2 May 2026

i keep thinking i want the models to be cheaper/faster more than i want them to be smarter but it seems that just being smarter is still the…

IndustryDGX agent

Sam Altman reflects on the trade-off between model performance improvements and practical considerations like cost and speed, suggesting that increased intelligence remains a primary focus despite use

RTX 5080 with 16 GB VRAM, 64 GB RAM best quantized model for programming?

Local AiDGX agent

For programming tasks with an RTX 5080 (16GB VRAM) and 64GB RAM, optimal quantized models include Qwen 3 14B at Q6 quantization, Llama 3.1 13B at Q8, or DeepSeek R1 Distill 14B Q4, all of which fit co

1 May 2026

A Grid-Aware Agent-Based Model for Analyzing Electric Vehicle Charging Systems

AgentsDGX agent

arXiv:2604.27849v1 Announce Type: new Abstract: This paper presents a configurable, grid-aware Agent-Based Model (ABM) for the systematic analysis of electric vehicle (EV) charging systems under confi

AI Models for Depressive Disorder Detection and Diagnosis: A Review

SafetyDGX agent

arXiv:2508.12022v2 Announce Type: replace Abstract: Major Depressive Disorder is one of the leading causes of disability worldwide, yet its diagnosis still depends largely on subjective clinical asses

Backdoor Attacks on Prompt-Driven Video Segmentation Foundation Models

AgentsDGX agent

arXiv:2512.22046v2 Announce Type: replace Abstract: Prompt-driven Video Segmentation Foundation Models (VSFMs), such as SAM2, are increasingly used in applications including autonomous driving and dig

Beyond the Mean: Within-Model Reliable Change Detection for LLM Evaluation

Model ReleasesDGX agent

arXiv:2604.27405v1 Announce Type: cross Abstract: We adapted the Reliable Change Index (RCI; Jacobson and Truax, 1991) from clinical psychology to item-level LLM version comparison on 2,000 MMLU-Pro i

Budget-Constrained Online Retrieval-Augmented Generation: The Chunk-as-a-Service Model

ResearchDGX agent

arXiv:2604.26981v1 Announce Type: cross Abstract: Large Language Models (LLMs) have revolutionized the field of natural language processing. However, they exhibit some limitations, including a lack of

CasLayout: Cascaded 3D Layout Diffusion for Indoor Scene Synthesis with Implicit Relation Modeling

Local AiDGX agent

arXiv:2604.27361v1 Announce Type: new Abstract: Synthesizing realistic 3D indoor scenes remains challenging due to data scarcity and the difficulty of simultaneously enforcing global architectural con

Cross-Lingual Response Consistency in Large Language Models: An ILR-Informed Evaluation of Claude Across Six Languages

Model ReleasesDGX agent

arXiv:2604.27137v1 Announce Type: new Abstract: This paper introduces a systematic evaluation framework grounded in the Interagency Language Roundtable (ILR) Skill Level Descriptions and applies it to

Detecting is Easy, Adapting is Hard: Local Expert Growth for Visual Model-Based Reinforcement Learning under Distribution Shift

Local AiDGX agent

arXiv:2604.27411v1 Announce Type: new Abstract: Visual model-based reinforcement learning (MBRL) agents can perform well on the training distribution, but often break down once the test environment sh

Efficient Sparse Selective-Update RNNs for Long-Range Sequence Modeling

TutorialsDGX agent

arXiv:2603.02226v2 Announce Type: replace Abstract: Real-world sequential signals, such as audio or video, contain critical information that is often embedded within long periods of silence or noise.

Exploring Applications of Transfer-State Large Language Models: Cognitive Profiling and Socratic AI Tutoring

SafetyDGX agent

arXiv:2604.27454v1 Announce Type: new Abstract: Large language models (LLMs) sometimes exhibit qualitative shifts in response style under sustained self-referential dialogue conditions (Berg et al., 2

Learning to Spend: Model Predictive Control for Budgeting under Non-Stationary Returns

ResearchDGX agent

arXiv:2604.27186v1 Announce Type: cross Abstract: We study finite-horizon budget allocation as a closed-loop economic control problem and evaluate receding-horizon Model Predictive Control (MPC) relat

Predicting Upcoming Stuttering Events from Three-Second Audio: Stratified Evaluation Reveals Severity-Selective Precursors, and the Model Deploys Fully On-Device

Model ReleasesDGX agent

arXiv:2604.27279v1 Announce Type: cross Abstract: Audio-based stuttering systems to date have been trained for detection -- what disfluency is present now -- leaving prediction, the capability needed

Proactive Dialogue Model with Intent Prediction

ResearchDGX agent

arXiv:2604.27379v1 Announce Type: new Abstract: Dialogue models are inherently reactive, responding to the current user turn without anticipating upcoming intents, which leads to redundant interaction

Repetition over Diversity: High-Signal Data Filtering for Sample-Efficient German Language Modeling

ResearchDGX agent

arXiv:2604.28075v1 Announce Type: cross Abstract: Recent research has shown that filtering massive English web corpora into high-quality subsets significantly improves training efficiency. However, fo

Sampler-Robust Optimization under Generative Models

ResearchDGX agent

arXiv:2604.27447v1 Announce Type: cross Abstract: Modern stochastic optimization pipelines increasingly rely on learned generative models to represent uncertainty, while downstream decisions are evalu

Semantic Structure of Feature Space in Large Language Models

ResearchDGX agent

arXiv:2604.27169v1 Announce Type: new Abstract: We show that the geometric relations between semantic features in large language models' hidden states closely mirror human psychological associations.

Simulating clinical interventions with a generative multimodal model of human physiology

ResearchDGX agent

arXiv:2604.27899v1 Announce Type: new Abstract: Understanding how human health changes over time, and why responses to interventions vary between individuals, remains a central challenge in medicine.

The TEA Nets framework combines AI and cognitive network science to model targets, events and actors in text

Model ReleasesDGX agent

arXiv:2604.27673v1 Announce Type: new Abstract: We introduce Target-Event-Agent Networks (TEA Nets) as a computational framework to extract subjects (``Agents'), verbs (``Events'), and objects (``Targ

Vision-Language Models Mistake Head Orientation for Gaze Direction: Nonverbal Conversation Cues

SafetyDGX agent

arXiv:2506.05412v3 Announce Type: replace-cross Abstract: Where someone looks is a nonverbal communication cue that children and adults readily use. How well can Vision-Language Models (VLMs) infer ga

You don't have to choose between either. It's best to use a combination of them. My advice is to learn how to use a few of these models in d…

TutorialsDGX agent

You don't have to choose between either. It's best to use a combination of them. My advice is to learn how to use a few of these models in different harnesses. Learn to combine their strengths. Open-w

30 Apr 2026

A Dual-Task Paradigm to Investigate Sentence Comprehension Strategies in Language Models

ResearchDGX agent

arXiv:2604.26351v1 Announce Type: new Abstract: Language models (LMs) behave more like humans when their cognitive resources are restricted, particularly in predicting sentence processing costs such a

ComboStoc: Combinatorial Stochasticity for Diffusion Generative Models

ResearchDGX agent

arXiv:2405.13729v3 Announce Type: replace-cross Abstract: In this paper, we study an under-explored but important factor of diffusion generative models, i.e., the combinatorial complexity. Data sample

Delta Score Matters! Spatial Adaptive Multi Guidance in Diffusion Models

SafetyDGX agent

arXiv:2604.26503v1 Announce Type: new Abstract: Diffusion models have achieved remarkable success in synthesizing complex static and temporal visuals, a breakthrough largely driven by Classifier-Free

DORA: A Scalable Asynchronous Reinforcement Learning System for Language Model Training

SafetyDGX agent

arXiv:2604.26256v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a critical paradigm for LLM post-training, yet the rollout phase -- accounting for 50--80% of total step time --

From Prompt Risk to Response Risk: Paired Analysis of Safety Behavior of Large Language Model

SafetyDGX agent

arXiv:2604.26052v1 Announce Type: new Abstract: Safety evaluations of large language models (LLMs) typically report binary outcomes such as attack success rate, refusal rate, or harmful/not-harmful re

HIVE: Hidden-Evidence Verification for Hallucination Detection in Diffusion Large Language Models

ResearchDGX agent

arXiv:2604.26139v1 Announce Type: new Abstract: Diffusion large language models generate text through multi-step denoising, where hallucination signals may emerge throughout the trajectory rather than

← Previous
1…189190191192193…1010
Next →