AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,522 results
Model Releases

When AI Tells You What You Want to Hear: Sycophantic Behavior of Large Language Models in Dementia Care Settings

DGX agent

arXiv:2605.16288v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in clinical and care settings. This exploratory study investigates whether LLMs exhibit sycophantic

model-releasesarxiv-cs-cl
19 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

A Model Can Help Itself: Reward-Free Self-Training for LLM Reasoning

DGX agent

arXiv:2510.18814v3 Announce Type: replace-cross Abstract: Can language models improve their reasoning performance without external rewards, using only their own sampled responses for training? We show

researcharxiv-cs-ai
18 May 2026
Model Releases

AGC: Adaptive Geodesic Correction for Adversarial Robustness on Vision-Language Models

DGX agent

arXiv:2605.15584v1 Announce Type: new Abstract: Vision-language models like CLIP have demonstrated remarkable zero-shot transfer capabilities. However, their susceptibility to imperceptible adversaria

model-releasesarxiv-cs-cv
18 May 2026
Research

Brain-OF: An Omnifunctional Foundation Model for fMRI, EEG and MEG

DGX agent

arXiv:2602.23410v3 Announce Type: replace-cross Abstract: Brain foundation models have achieved remarkable advances across a wide range of neuroscience tasks. However, most existing models are limited

researcharxiv-cs-ai
18 May 2026
Model Releases

CLARE: Continual Learning for Vision-Language-Action Models via Autonomous Adapter Routing and Expansion

DGX agent

arXiv:2601.09512v2 Announce Type: replace-cross Abstract: To teach robots complex manipulation tasks, a common approach is to fine-tune a pre-trained vision-language-action model (VLA) on task-specifi

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Entity-Centric World Models: Interaction-Aware Masking for Causal Video Prediction

DGX agent

arXiv:2605.15466v1 Announce Type: new Abstract: Learning predictive world models from unlabelled video is a foundational challenge in artificial intelligence. While Joint Embedding Predictive Architec

model-releasesarxiv-cs-cv
18 May 2026
Safety

EntropyScan: Towards Model-level Backdoor Detection in LVLMs via Visual Attention Entropy

DGX agent

arXiv:2605.15711v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have demonstrated remarkable capabilities across various tasks, yet they remain vulnerable to backdoor attacks. Exi

safetyarxiv-cs-cv
18 May 2026
Applications

Going Beyond the Edge: Distributed Inference of Transformer Models on Ultra-Low-Power Wireless Devices

DGX agent

arXiv:2605.15694v1 Announce Type: new Abstract: Transformer models are rapidly becoming a cornerstone of modern Internet of Things (IoT) applications, yet their computational and memory demands far ex

applicationsarxiv-cs-lg
18 May 2026
Safety

Improve Large Language Model Systems with User Logs

DGX agent

arXiv:2602.06470v2 Announce Type: replace-cross Abstract: Scaling training data and model parameters has long driven progress in large language models (LLMs), but this paradigm is increasingly constra

safetyarxiv-cs-ai
18 May 2026
Research

Neural Activation Patterns Across Language Model Architectures: A Comprehensive Analysis of Cognitive Task Performance

DGX agent

arXiv:2605.15436v1 Announce Type: new Abstract: This paper presents a comprehensive analysis of neural activation patterns across six distinct large language model (LLM) architectures, examining their

researcharxiv-cs-cl
18 May 2026
Model Releases

Painless Activation Steering: An Automated, Lightweight Approach for Post-Training Large Language Models

DGX agent

arXiv:2509.22739v3 Announce Type: replace-cross Abstract: Language models (LMs) are typically post-trained for desired capabilities and behaviors via weight-based or prompt-based steering, but the for

model-releasesarxiv-cs-ai
18 May 2026
Research

Privacy Evaluation of Generative Models for Trajectory Generation

DGX agent

arXiv:2605.15246v1 Announce Type: new Abstract: Trajectory data is fundamental to modern urban intelligence, yet its sensitivity raises significant privacy concerns. Generative models such as Generati

researcharxiv-cs-lg
18 May 2026
Research

RoiMAM: Region-of-Interest Medical Attention Model for Efficient Vision-Language Understanding

DGX agent

arXiv:2605.15561v1 Announce Type: new Abstract: Vision-Language Models (VLMs) facilitate medical visual question answering (MedVQA) by jointly interpreting images and text. However, existing models ty

researcharxiv-cs-cv
18 May 2026
Tutorials

Simultaneous State Estimation and Online Model Learning in a Soft Robotic System

DGX agent

arXiv:2602.14092v2 Announce Type: replace-cross Abstract: Operating complex real-world systems, such as soft robots, can benefit from precise predictive control schemes that require accurate state and

tutorialsarxiv-cs-ro
18 May 2026
Model Releases

Steve Bannon and 60+ Trump allies sign a Humans First-led letter urging Trump to mandate government testing and approval of powerful AI models before release (Ashley Gold/Axios)

DGX agent

Ashley Gold / Axios: Steve Bannon and 60+ Trump allies sign a Humans First-led letter urging Trump to mandate government testing and approval of powerful AI models before release — A group of more tha

model-releasestechmeme
18 May 2026
Applications

Toward World Modeling of Physiological Signals with Chaos-Theoretic Balancing and Latent Dynamics

DGX agent

arXiv:2605.15465v1 Announce Type: new Abstract: Physiological time series signals reflect complex, multi-scale dynamical processes of the human body. Existing modeling studies focus on static tasks su

applicationsarxiv-cs-lg
18 May 2026
Agents

Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk Prediction

DGX agent

arXiv:2510.10454v2 Announce Type: replace Abstract: Large language models (LLMs) offer a generalizable approach for modeling patient trajectories, but suffer from the long and noisy nature of electron

agentsarxiv-cs-ai
18 May 2026
Model Releases

Transformer Scalability Crisis: The First Comprehensive Empirical Analysis of Performance Walls in Modern Language Models

DGX agent

arXiv:2605.15413v1 Announce Type: new Abstract: Despite the remarkable success of transformer architectures in natural language processing, their scalability limitations remain poorly understood throu

model-releasesarxiv-cs-lg
18 May 2026
Industry

Very cool to see Cursor doubling down on training great models. In my opinion, ultimately all serious companies in AI will want to train mod…

DGX agent

Very cool to see Cursor doubling down on training great models. In my opinion, ultimately all serious companies in AI will want to train models themselves, based on open-source instead of outsourcing

industryclem-delangue--x
18 May 2026
Safety

Video Models Can Reason with Verifiable Rewards

DGX agent

arXiv:2605.15458v1 Announce Type: new Abstract: Video diffusion models have made rapid progress in perceptual realism and temporal coherence, but they remain primarily optimized for plausible generati

safetyarxiv-cs-cv
18 May 2026
Model Releases

When do you reach for other models instead of Claude? What can we do better? Hit me with all of your frustrations. dms open. If you can give…

DGX agent

When do you reach for other models instead of Claude? What can we do better? Hit me with all of your frustrations. dms open. If you can give me detail (e.g. specifics/transcipts) - it'll help a lot in

model-releasesthariq--x
17 May 2026
Model Releases

AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models

DGX agent

arXiv:2509.26100v2 Announce Type: replace Abstract: The rapid integration of Large Language Models (LLMs) into high-stakes domains necessitates reliable safety and compliance evaluation. However, exis

model-releasesarxiv-cs-ai
15 May 2026
Safety

Agentifying Patient Dynamics within LLMs through Interacting with Clinical World Model

DGX agent

arXiv:2605.14723v1 Announce Type: new Abstract: Sepsis management in the ICU requires sequential treatment decisions under rapidly evolving patient physiology. Although large language models (LLMs) en

safetyarxiv-cs-ai
15 May 2026
Model Releases

CounselBench: A Large-Scale Expert Evaluation and Adversarial Benchmarking of Large Language Models in Mental Health Question Answering

DGX agent

arXiv:2506.08584v4 Announce Type: replace Abstract: Medical question answering (QA) benchmarks often focus on multiple-choice or fact-based tasks, leaving open-ended answers to real patient questions

model-releasesarxiv-cs-cl
15 May 2026
Applications

Energy-Regularized Sequential Model Editing on Hyperspheres

DGX agent

arXiv:2510.01172v3 Announce Type: replace Abstract: Large language models (LLMs) require constant updates to remain aligned with evolving real-world knowledge. Model editing offers a lightweight alter

applicationsarxiv-cs-cl
15 May 2026
Model Releases

Exploring Vision-Language Models for Online Signature Verification: A Zero-Shot Capability Study

DGX agent

arXiv:2605.14845v1 Announce Type: new Abstract: Recent advancements in Vision-Language Models (VLMs) have demonstrated strong capabilities in general visual reasoning, yet their applicability to rigor

model-releasesarxiv-cs-cv
15 May 2026
Research

Generative Bayesian Optimization: Generative Models as Acquisition Functions

DGX agent

arXiv:2510.25240v3 Announce Type: replace-cross Abstract: We present a general strategy for turning generative models into candidate solution samplers for batch Bayesian optimization (BO). The use of

researcharxiv-cs-lg
15 May 2026
Model Releases

Hyperspectral Image Land Cover Captioning Dataset for Vision Language Models

DGX agent

arXiv:2505.12217v2 Announce Type: replace Abstract: We introduce HyperCap, the first large-scale hyperspectral captioning dataset designed to enhance model performance and effectiveness in remote sens

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Kairos: Toward Adaptive and Parameter-Efficient Time Series Foundation Models

DGX agent

arXiv:2509.25826v3 Announce Type: replace Abstract: Inherent temporal heterogeneity, such as varying sampling densities and periodic structures, has posed substantial challenges in zero-shot generaliz

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Mechanistic Interpretability of EEG Foundation Models via Sparse Autoencoders

DGX agent

arXiv:2605.13930v1 Announce Type: new Abstract: EEG foundation models achieve state-of-the-art clinical performance, yet the internal computations driving their predictions remain opaque: a barrier to

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

MultiEmo-Bench: Multi-label Visual Emotion Analysis for Multi-modal Large Language Models

DGX agent

arXiv:2605.14635v1 Announce Type: cross Abstract: This paper introduces a multi-label visual emotion analysis benchmark dataset for comprehensively evaluating the ability of multimodal large language

model-releasesarxiv-cs-ai
15 May 2026
Agents

Orchard: An Open-Source Agentic Modeling Framework

DGX agent

arXiv:2605.15040v1 Announce Type: new Abstract: Agentic modeling aims to transform LLMs into autonomous agents capable of solving complex tasks through planning, reasoning, tool use, and multi-turn in

agentsarxiv-cs-ai
15 May 2026
Local Ai

Small, Private Language Models as Teammates for Educational Assessment Design

DGX agent

arXiv:2605.15015v1 Announce Type: new Abstract: Generative AI increasingly supports educational design tasks, e.g., through Large Language Models (LLMs), demonstrating the capability to design assessm

local-aiarxiv-cs-ai
15 May 2026
Model Releases

Three new open-source models just landed in ComfyUI natively: → Gemma 4 (Google DeepMind) - multimodal LLM handling text, image, audio, and …

DGX agent

Three new open-source models just landed in ComfyUI natively: → Gemma 4 (Google DeepMind) - multimodal LLM handling text, image, audio, and video input with built-in step-by-step reasoning mode → VOID

model-releasescomfyui--x
15 May 2026
Safety

XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations

DGX agent

arXiv:2511.02776v2 Announce Type: replace Abstract: Recent progress in large-scale robotic datasets and vision-language models (VLMs) has advanced research on vision-language-action (VLA) models. Howe

safetyarxiv-cs-ro
15 May 2026
Research

Amortized Guidance for Image Inpainting with Pretrained Diffusion Models

DGX agent

arXiv:2605.13010v1 Announce Type: cross Abstract: We study image inpainting with generative diffusion models. Existing methods typically either train dedicated task-specific models, or adapt a pretrai

researcharxiv-cs-ai
14 May 2026
Model Releases

Bias In, Bias Out? Finding Unbiased Subnetworks in Vanilla Models

DGX agent

arXiv:2603.05582v2 Announce Type: replace-cross Abstract: The issue of algorithmic biases in deep learning has led to the development of various debiasing techniques, many of which perform complex tra

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Continual Fine-Tuning of Large Language Models via Program Memory

DGX agent

arXiv:2605.13162v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT), particularly Low-Rank Adaptation (LoRA), has become a standard approach for adapting Large Language Models (LLMs

model-releasesarxiv-cs-lg
14 May 2026
Agents

Earth Science Foundation Models: From Perception to Reasoning and Discovery

DGX agent

arXiv:2605.12542v1 Announce Type: cross Abstract: Large foundation models (FMs) are transforming Earth science by integrating heterogeneous multimodal data, such as multi-platform imagery, gridded rea

agentsarxiv-cs-lg
14 May 2026
Model Releases

Lifelong Learning in Vision-Language Models: Enhanced EWC with Cross-Modal Knowledge Retention

DGX agent

arXiv:2605.12789v1 Announce Type: new Abstract: Large language-vision models (LVLMs) such as CLIP, Flamingo, and BLIP have revolutionized AI by enabling understanding across textual and visual modalit

model-releasesarxiv-cs-ro
14 May 2026
Research

Prismatic World Model: Learning Compositional Dynamics for Planning in Hybrid Systems

DGX agent

arXiv:2512.08411v2 Announce Type: replace Abstract: Model-based planning in robotic domains is challenged by the hybrid nature of physical dynamics, where continuous motion is punctuated by discrete e

researcharxiv-cs-ai
14 May 2026
Applications

RotVLA: Rotational Latent Action for Vision-Language-Action Model

DGX agent

arXiv:2605.13403v1 Announce Type: cross Abstract: Latent Action Models (LAMs) have emerged as an effective paradigm for handling heterogeneous datasets during Vision-Language-Action (VLA) model pretra

applicationsarxiv-cs-cv
14 May 2026
Research

Sampling from Flow Language Models via Marginal-Conditioned Bridges

DGX agent

arXiv:2605.13681v1 Announce Type: new Abstract: Flow Language Models (FLMs) are a recently introduced class of language models which adapt continuous flow matching for one-hot encoded token sequences.

researcharxiv-cs-lg
14 May 2026
Research

Three-Stage Learning Unlocks Strong Performance in Simple Models for Long-Term Time Series Forecasting

DGX agent

arXiv:2605.13678v1 Announce Type: new Abstract: Recent studies on long-term time series forecasting have shown that simple linear models and MLP-based predictors can achieve strong performance without

researcharxiv-cs-lg
14 May 2026
Model Releases

Towards Long-horizon Embodied Agents with Tool-Aligned Vision-Language-Action Models

DGX agent

arXiv:2605.13119v1 Announce Type: cross Abstract: Vision-language-action (VLA) models are effective robot action executors, but they remain limited on long-horizon tasks due to the dual burden of exte

model-releasesarxiv-cs-ai
14 May 2026
Agents

Training Long-Context Vision-Language Models Effectively with Generalization Beyond 128K Context

DGX agent

arXiv:2605.13831v1 Announce Type: new Abstract: Long-context modeling is becoming a core capability of modern large vision-language models (LVLMs), enabling sustained context management across long-do

agentsarxiv-cs-cv
14 May 2026
Research

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone

DGX agent

arXiv:2605.11405v1 Announce Type: new Abstract: Data curation has shifted the quality-compute frontier for language-model and contrastive image-text pretraining, but its role for vision-language model

researcharxiv-cs-lg
13 May 2026
Model Releases

A Proof-of-Concept Simulation-Driven Digital Twin Framework for Decision-Aware Diabetes Modeling

DGX agent

arXiv:2605.11247v1 Announce Type: new Abstract: This paper presents a proof-of-concept digital twin framework for simulation-driven diabetes modeling using benchmark clinical data, synthetic temporal

model-releasesarxiv-cs-lg
13 May 2026
← Previous
1…111112113114115…1261
Next →