AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,565 results
6 May 2026

CellxPert: Inference-Time MCMC Steering of a Multi-Omics Single-Cell Foundation Model for In-Silico Perturbation

ResearchDGX agent

arXiv:2605.00930v1 Announce Type: cross Abstract: In this work, we introduce CellxPert, a scalable multimodal foundation model that unifies single-cell and spatial multi-omics within a common represen

Dual-Foundation Models for Unsupervised Domain Adaptation

AgentsDGX agent

arXiv:2605.03365v1 Announce Type: new Abstract: Semantic segmentation provides pixel-level scene understanding essential for autonomous driving and fine-grained perception tasks. However, training seg

FINER-SQL: Boosting Small Language Models for Text-to-SQL

SafetyDGX agent

arXiv:2605.03465v1 Announce Type: cross Abstract: Large language models have driven major advances in Text-to-SQL generation. However, they suffer from high computational cost, long latency, and data

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

From Experimental Limits to Physical Insight: A Retrieval-Augmented Multi-Agent Framework for Interpreting Searches Beyond the Standard Model

AgentsDGX agent

arXiv:2605.02491v1 Announce Type: cross Abstract: Modern searches for physics beyond the Standard Model produce rapidly expanding literature containing heterogeneous information, including textual ana

Google, Microsoft and xAI agree to allow government safety checks of their AI models prior to release

SafetyDGX agent

Google LLC, Microsoft Corp. and xAI have agreed to share unreleased versions of their artificial intelligence models with the U.S. Department of Commerce to ensure the technologies do not pose a threa

Grounding Multi-Hop Reasoning in Structural Causal Models via Group Relative Policy Optimization

SafetyDGX agent

arXiv:2605.01482v1 Announce Type: new Abstract: Multi-Hop Fact Verification (MHFV) necessitates complex reasoning across disparate evidence, posing significant challenges for Large Language Models (LL

Memory-Efficient Continual Learning with CLIP Models

ResearchDGX agent

arXiv:2605.03866v1 Announce Type: new Abstract: Contrastive Language-Image Pretraining (CLIP) models excel at understanding image-text relationships but struggle with adapting to new data without forg

Model Routing as a Trust Problem: Route Receipts for Adaptive AI Systems

SafetyDGX agent

arXiv:2605.01710v1 Announce Type: new Abstract: AI products often route requests through version aliases, service tiers, tool choices, regional endpoints, fallback rules, or safety handling before res

Partially Observed Structural Causal Models

ResearchDGX agent

arXiv:2605.03268v1 Announce Type: new Abstract: Here we introduce Partially Observed Structural Causal Models (POSCMs) that formalize causal systems where latent contexts co-determine both the interac

Personalized Digital Health Modeling with Adaptive Support Users

ApplicationsDGX agent

arXiv:2605.02004v1 Announce Type: new Abstract: Personalized models are essential in digital health because individuals exhibit substantial physiological and behavioral heterogeneity. Yet personalizat

Prism: Efficient Test-Time Scaling via Hierarchical Search and Self-Verification for Discrete Diffusion Language Models

Model ReleasesDGX agent

arXiv:2602.01842v3 Announce Type: replace Abstract: Inference-time compute has re-emerged as a practical way to improve LLM reasoning. Most test-time scaling (TTS) algorithms rely on autoregressive de

Proteo-R1: Reasoning Foundation Models for De Novo Protein Design

ResearchDGX agent

arXiv:2605.02937v1 Announce Type: new Abstract: Deep learning in de novo protein design has achieved atomic-level fidelity. However, existing models remain largely non-deliberative: they directly synt

Sparse Data Tree Canopy Segmentation: Fine-Tuning Leading Pretrained Models on Only 150 Images

ResearchDGX agent

arXiv:2601.10931v2 Announce Type: replace Abstract: Tree canopy detection from aerial imagery is an important task for environmental monitoring, urban planning, and ecosystem analysis. Simulating real

When Safety Geometry Collapses: Fine-Tuning Vulnerabilities in Agentic Guard Models

SafetyDGX agent

arXiv:2605.02914v1 Announce Type: new Abstract: A guard model fine-tuned on entirely benign data can lose all safety alignment -- not through adversarial manipulation, but through standard domain spec

5 May 2026

CADFS: A Big CAD Program Dataset and Framework for Computer-Aided Design with Large Language Models

ApplicationsDGX agent

arXiv:2605.01925v1 Announce Type: new Abstract: We introduce CADFS, a data-centric framework that enables large vision-language models to generate complex CAD design histories. Existing generative CAD

Compression as Adaptation: Implicit Visual Representation with Diffusion Foundation Models

ResearchDGX agent

arXiv:2603.07615v2 Announce Type: replace-cross Abstract: Modern visual generative models acquire rich visual knowledge through large-scale training, yet existing visual representations (such as pixel

Do Large Language Models Plan Answer Positions? Position Bias in Multiple-Choice Question Generation

SafetyDGX agent

arXiv:2605.01846v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to generate multiple-choice questions (MCQs), where correct answers should ideally be uniformly distr

Edge-Efficient Image Restoration: Transformer Distillation into State-Space Models

ResearchDGX agent

arXiv:2605.02794v1 Announce Type: new Abstract: We propose a modular framework for hybrid image restoration that integrates transformer and state-space model (SSM) blocks with a focus on improving run

Extending machine learning model for implicit solvation to free energy calculations

ResearchDGX agent

arXiv:2510.20103v2 Announce Type: replace-cross Abstract: The implicit solvent approach offers a computationally efficient framework to model solvation effects in molecular simulations. However, its a

Extreme Weather Bench: A framework and benchmark for evaluation of high-impact weather

Model ReleasesDGX agent

arXiv:2605.01126v1 Announce Type: new Abstract: Forecasting the wide variety of high-impact weather events experienced globally is a challenge for both Artificial Intelligence (AI) and Numerical Weath

Fuzzy Fingerprinting Encoder Pre-trained Language Models for Emotion Recognition in Conversations: Human Assessment and Validity Study

ResearchDGX agent

arXiv:2605.02665v1 Announce Type: new Abstract: In Emotion Recognition in Conversations (ERC), model decisions should align with nuanced human perception and ideally provide insights on the classifica

GEASS: Training-Free Caption Steering for Hallucination Mitigation in Vision-Language Models

ResearchDGX agent

arXiv:2605.01733v1 Announce Type: new Abstract: Vision-Language Models (VLMs) excel at grounded reasoning but remain prone to object hallucination. Recent work treats self-generated captions as a unif

GeoSAE: Geometric Prior-Guided Layer-Wise Sparse Autoencoder Annotation of Brain MRI Foundation Models

TutorialsDGX agent

arXiv:2605.01829v1 Announce Type: new Abstract: Brain MRI foundation models learn rich representations of anatomy, but interpreting what clinical information they encode remains an open problem. Stand

H-Probes: Extracting Hierarchical Structures From Latent Representations of Language Models

ApplicationsDGX agent

arXiv:2605.00847v1 Announce Type: new Abstract: Representing and navigating hierarchy is a fundamental primitive of reasoning. Large language models have demonstrated proficiency in a wide variety of

How Prompts Move Language Model Behavior: Frames, Salience, and Construal as Semantic Control

ResearchDGX agent

arXiv:2512.12688v3 Announce Type: replace-cross Abstract: Prompt engineering is widely used to shape large language model behavior, yet it is often treated as a practical heuristic rather than as a fo

InfoLaw: Information Scaling Laws for Large Language Models with Quality-Weighted Mixture Data and Repetition

ResearchDGX agent

arXiv:2605.02364v1 Announce Type: new Abstract: Upweighting high-quality data in LLM pretraining often improves performance, but in datalimited regimes, especially under overtraining, stronger upweigh

LLM-Foraging: Large Language Models for Decentralized Swarm Robot Foraging

Model ReleasesDGX agent

arXiv:2605.01461v1 Announce Type: new Abstract: Swarm foraging algorithms, such as the central-place foraging algorithm (CPFA), typically rely on offline parameter optimization using genetic algorithm

Model-Based Proactive Cost Generation for Learning Safe Policies Offline with Limited Violation Data

SafetyDGX agent

arXiv:2605.01356v1 Announce Type: new Abstract: Learning constraint-satisfying policies from offline data without risky online interaction is crucial for safety-critical decision making. Conventional

Noise is All You Need: Solving Linear Inverse Problems by Noise Combination Sampling with Diffusion Models

ResearchDGX agent

arXiv:2510.23633v2 Announce Type: replace-cross Abstract: Pretrained diffusion models have demonstrated strong capabilities in zero-shot inverse problem solving by incorporating observation informatio

open-weight LLMs have come a long way on agent tasks! but the harness you wrap them in matters just as much as the model itself, and arguabl…

AgentsDGX agent

open-weight LLMs have come a long way on agent tasks! but the harness you wrap them in matters just as much as the model itself, and arguably the interface you use to drive that harness matters even m

OpenAI GPT-5 System Card

Model ReleasesDGX agent

arXiv:2601.03267v2 Announce Type: replace Abstract: This is the system card published alongside the OpenAI GPT-5 launch, August 2025. GPT-5 is a unified system with a smart and fast model that answers

PACE: Post-Causal Entropy Modeling for Learned LiDAR Point Cloud Compression

AgentsDGX agent

arXiv:2605.01320v1 Announce Type: new Abstract: LiDAR point cloud compression is vital for autonomous systems to handle massive data from high-resolution sensors. While learned entropy modeling built

Parllama -- a terminal UI for Ollama model management and multi-provider LLM chat

Local AiDGX agent

Parllama is a TUI (Text UI) application designed for easy management and use of Ollama-based LLMs that also works with major cloud-provided LLMs. It provides core model management features including f

PC-MNet: Dual-Level Congruity Modeling for Multimodal Sarcasm Detection via Polarity-Modulated Attention

Model ReleasesDGX agent

arXiv:2605.02447v1 Announce Type: new Abstract: Multimodal sarcasm detection, which aims to precisely identify pragmatic incongruities between literal text and nonverbal cues, has gained substantial a

PubMed-Ophtha: An open resource for training ophthalmology vision-language models on scientific literature

ResearchDGX agent

arXiv:2605.02720v1 Announce Type: cross Abstract: Vision-language models hold considerable promise for ophthalmology, but their development depends on large-scale, high-quality image-text datasets tha

Rethinking Model Selection in VLM Through the Lens of Gromov-Wasserstein Distance

SafetyDGX agent

arXiv:2605.01325v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have enhanced traditional LLMs with visual capabilities through the integration of vision encoders. While recent works hav

Scaling Vision Transformers for Functional MRI with Flat Maps

Model ReleasesDGX agent

arXiv:2510.13768v2 Announce Type: replace Abstract: We study the problem of training self-supervised foundation models for functional MRI. Our main contributions are: (1) we introduce a new model fami

Spatiotemporal Hidden-State Dynamics as a Signature of Internal Reasoning in Large Language Models

ResearchDGX agent

arXiv:2605.01853v1 Announce Type: new Abstract: Large reasoning models (LRMs) generate extended solutions, yet it remains unclear whether these traces reflect substantive internal computation or merel

SpectraDINO: Bridging the Spectral Gap in Vision Foundation Models via Lightweight Adapters

SafetyDGX agent

arXiv:2605.02258v1 Announce Type: new Abstract: Vision Foundation Models (VFMs) pretrained on large-scale RGB data have demonstrated remarkable representation quality, yet their applicability to multi

Spoken Language Identification with Pre-trained Models and Margin Loss

Model ReleasesDGX agent

arXiv:2605.01905v1 Announce Type: cross Abstract: For the speaker-controlled spoken language identification task proposed in the TidyLang Challenge 2026, this paper proposes a language identification

Stochastic Modeling of Human-Machine Authentication Channels under Partial Information Leakage

ApplicationsDGX agent

arXiv:2605.02102v1 Announce Type: cross Abstract: Reliable and secure human-machine communication is fundamental to IoT and cyber-physical ecosystems, where smartphones and wearables commonly serve as

StressEval: Failure-Driven Dynamic Benchmarking for Knowledge-Intensive Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2605.01939v1 Announce Type: new Abstract: Static benchmarks for LLMs are increasingly compromised by contamination and overfitting especially on knowledge intensive reasoning tasks While recent

Stylistic Attribute Control in Latent Diffusion Models

ResearchDGX agent

arXiv:2605.02583v1 Announce Type: new Abstract: Text-to-image diffusion models have revolutionized image synthesis and editing, but precise control over stylistic attributes remains a challenge, often

VAnim: Rendering-Aware Sparse State Modeling for Structure-Preserving Vector Animation

Model ReleasesDGX agent

arXiv:2605.01517v1 Announce Type: new Abstract: Scalable Vector Graphics (SVG) animation generation is pivotal for professional design due to their structural editability and resolution independence.

When To Adapt? Adapting the Model or Data in Federated Medical Imaging

ResearchDGX agent

arXiv:2605.00892v1 Announce Type: new Abstract: Federated learning enables collaborative model training across medical institutions without sharing raw data, but its performance is often limited by do

4 May 2026

Federated Weather Modeling on Sensor Data

ResearchDGX agent

arXiv:2605.00322v1 Announce Type: new Abstract: Federated weather modeling on sensor data is a distributed system underpinned by federated learning, enabling multiple sensor data sources, including gr

Intelligent Elastic Feature Fading: Enabling Model Retrain-Free Feature Efficiency Rollouts at Scale

SafetyDGX agent

arXiv:2605.00324v1 Announce Type: cross Abstract: Large-scale ranking systems depend on thousands of features derived from user behavior across multiple time horizons. Typically requires model retrain

It's Never Too Late: Noise Optimization for Collapse Recovery in Trained Diffusion Models

ResearchDGX agent

arXiv:2601.00090v2 Announce Type: replace Abstract: Contemporary text-to-image models exhibit a surprising degree of mode collapse, as can be seen when sampling several images given the same text prom

NEW paper from Sakana AI (ICLR 2026). A 7B Conductor model just hit SOTA on GPQA-Diamond and LiveCodeBench by orchestrating other LLMs inste…

SafetyDGX agent

NEW paper from Sakana AI (ICLR 2026). A 7B Conductor model just hit SOTA on GPQA-Diamond and LiveCodeBench by orchestrating other LLMs instead of solving problems itself. (great paper! bookmark it!) T

Reward Modeling from Natural Language Human Feedback

ResearchDGX agent

arXiv:2601.07349v3 Announce Type: replace Abstract: Reinforcement Learning with Verifiable reward (RLVR) on preference data has become the mainstream approach for training Generative Reward Models (GR

When Do Diffusion Models learn to Generate Multiple Objects?

TutorialsDGX agent

arXiv:2605.00273v1 Announce Type: new Abstract: Text-to-image diffusion models achieve impressive visual fidelity, yet they remain unreliable in multi-object generation. Despite extensive empirical ev

3 May 2026

Best Local Vision-Language Models?

Local AiDGX agent

This discussion thread explores lightweight vision-language models that can run locally, including options like Llama 3.2 Vision, Qwen2.5-VL, and SmolVLM2, optimized for tasks like OCR and visual ques

1 May 2026

Aligning Perception, Reasoning, Modeling and Interaction: A Survey on Physical AI

ApplicationsDGX agent

arXiv:2510.04978v5 Announce Type: replace Abstract: The rapid advancement of embodied intelligence and world models has intensified efforts to integrate physical laws into AI systems, yet physical per

Beyond Semantics: Measuring Fine-Grained Emotion Preservation in Small Language Model-Based Machine Translation

Model ReleasesDGX agent

arXiv:2604.27920v1 Announce Type: cross Abstract: Preserving affective nuance remains a challenge in Machine Translation (MT), where semantic equivalence often takes precedence over emotional fidelity

Cross-Lingual Sentiment Misalignment: Auditing Multilingual Language Models for Inversion Risk, Dialectal Representation, and Affective Stability

SafetyDGX agent

arXiv:2602.17469v2 Announce Type: replace Abstract: Recent advances in multilingual representation learning aim to bridge the performance gap between high- and low-resource languages, yet their abilit

Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection

Model ReleasesDGX agent

arXiv:2501.04066v2 Announce Type: replace Abstract: As a special type of multimedia data, Lithography Hotspot Detection (LHD) training often requires stronger privacy protection than conventional mult

From Context to Skills: Can Language Models Learn from Context Skillfully?

AgentsDGX agent

arXiv:2604.27660v1 Announce Type: new Abstract: Many real-world tasks require language models (LMs) to reason over complex contexts that exceed their parametric knowledge. This calls for context learn

GourNet: A CNN-Based Model for Mango Leaf Disease Detection

ApplicationsDGX agent

arXiv:2604.27764v1 Announce Type: new Abstract: Mango cultivation is crucial in the agricultural sector, significantly contributing to economic development and food security. However, diseases affecti

Green Physics-Informed Machine Learning Models For Structural Health Monitoring

ApplicationsDGX agent

arXiv:2604.27638v1 Announce Type: new Abstract: Machine learning continues to emerge as an important tool to be utilised within structural engineering and structural health monitoring, due to its abil

Improving Calibration in Test-Time Prompt Tuning for Vision-Language Models via Data-Free Flatness-Aware Prompt Pretraining

ApplicationsDGX agent

arXiv:2604.27715v1 Announce Type: new Abstract: Test-time prompt tuning (TPT) has emerged as a promising technique for enhancing the adaptability of vision-language models by optimizing textual prompt

← Previous
1…162163164165166…1010
Next →