AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,522 results
22 Apr 2026

What’s new in Cloud Run at Next ‘26

Model ReleasesDGX agent

From vibe-coded and large-scale apps to AI models and agents, Cloud Run delivers on-demand compute with zero overhead and pay-per-use pricing for all of your workloads. Last year, the number of extern

21 Apr 2026

A Systematic Survey and Benchmark of Deep Learning for Molecular Property Prediction in the Foundation Model Era

Model ReleasesDGX agent

arXiv:2604.16586v1 Announce Type: new Abstract: Molecular property prediction integrates quantum chemistry, cheminformatics, and deep learning to connect molecular structure with physicochemical and b

A Unification of Discrete, Gaussian, and Simplicial Diffusion

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2512.15923v2 Announce Type: replace Abstract: To model discrete sequences such as DNA, proteins, and language using diffusion, practitioners must choose between three major methods: diffusion in

An LLM-Guided Query-Aware Inference System for GNN Models on Large Knowledge Graphs

ApplicationsDGX agent

arXiv:2603.04545v2 Announce Type: replace Abstract: Efficient inference for graph neural networks (GNNs) on large knowledge graphs (KGs) is essential for many real-world applications. GNN inference qu

BEFT: Bias-Efficient Fine-Tuning of Language Models in Low-Data Regimes

Model ReleasesDGX agent

arXiv:2509.15974v2 Announce Type: replace Abstract: Fine-tuning the bias terms of large language models (LLMs) has the potential to achieve unprecedented parameter efficiency while maintaining competi

Calibrating Model-Based Evaluation Metrics for Summarization

ResearchDGX agent

arXiv:2604.17200v1 Announce Type: new Abstract: Recent advances in summary evaluation are based on model-based metrics to assess quality dimensions, such as completeness, conciseness, and faithfulness

Closing the Modality Reasoning Gap for Speech Large Language Models

SafetyDGX agent

arXiv:2601.05543v2 Announce Type: replace Abstract: Although Speech Large Language Models have achieved notable progress, a substantial modality reasoning gap remains: their reasoning performance on s

ConMeZO: Adaptive Descent-Direction Sampling for Gradient-Free Finetuning of Large Language Models

Model ReleasesDGX agent

arXiv:2511.02757v2 Announce Type: replace Abstract: Zeroth-order or derivative-free optimization (MeZO) is an attractive strategy for finetuning large language models (LLMs) because it eliminates the

Domain-Specialized Object Detection via Model-Level Mixtures of Experts

ResearchDGX agent

arXiv:2604.18256v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models provide a structured approach to combining specialized neural networks and offer greater interpretability than conventio

Employing General-Purpose and Biomedical Large Language Models with Advanced Prompt Engineering for Pharmacoepidemiologic Study Design

Model ReleasesDGX agent

arXiv:2604.17988v1 Announce Type: new Abstract: Background: The potential of large language models (LLMs) to automate and support pharmacoepidemiologic study design is an emerging area of interest, ye

Erasing Thousands of Concepts: Towards Scalable and Practical Concept Erasure for Text-to-Image Diffusion Models

SafetyDGX agent

arXiv:2604.16481v1 Announce Type: new Abstract: Large-scale text-to-image (T2I) diffusion models deliver remarkable visual fidelity but pose safety risks due to their capacity to reproduce undesirable

ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection

Model ReleasesDGX agent

arXiv:2410.04509v3 Announce Type: replace Abstract: As the field of Multimodal Large Language Models (MLLMs) continues to evolve, their potential to revolutionize artificial intelligence is particular

Establishing a Scale for Kullback-Leibler Divergence in Language Models Across Various Settings

ResearchDGX agent

arXiv:2505.15353v3 Announce Type: replace Abstract: Log-likelihood vectors define a common space for comparing language models as probability distributions, enabling unified comparisons across heterog

FedLLM: A Privacy-Preserving Federated Large Language Model for Explainable Traffic Flow Prediction

Model ReleasesDGX agent

arXiv:2604.16612v1 Announce Type: new Abstract: Traffic prediction plays a central role in intelligent transportation systems (ITS) by supporting real-time decision-making, congestion management, and

FLARE: Task-agnostic embedding model evaluation through a normalization process

Model ReleasesDGX agent

arXiv:2604.17344v1 Announce Type: cross Abstract: When task-specific labels are not available, it becomes difficult to select an embedding model for a specific target corpus. Existing labelless measur

GaLa: Hypergraph-Guided Visual Language Models for Procedural Planning

ResearchDGX agent

arXiv:2604.17241v1 Announce Type: new Abstract: Implicit spatial relations and deep semantic structures encoded in object attributes are crucial for procedural planning in embodied AI systems. However

Graph neural network for colliding particles with an application to sea ice floe modeling

TutorialsDGX agent

arXiv:2602.16213v2 Announce Type: replace-cross Abstract: This paper introduces a novel approach to sea ice modeling using Graph Neural Networks (GNNs), utilizing the natural graph structure of sea ic

How Language Models Conflate Logical Validity with Plausibility: A Representational Analysis of Content Effects

SafetyDGX agent

arXiv:2510.06700v3 Announce Type: replace Abstract: Both humans and large language models (LLMs) exhibit content effects: biases in which the plausibility of the semantic content of a reasoning proble

How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study

Model ReleasesDGX agent

arXiv:2505.15404v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have achieved remarkable success on reasoning-intensive tasks such as mathematics and programming. However, their enha

How to Approximate Inference with Subtractive Mixture Models

TutorialsDGX agent

arXiv:2604.16714v1 Announce Type: new Abstract: Classical mixture models (MMs) are widely used tractable proposals for approximate inference settings such as variational inference (VI) and importance

Introducing ChatGPT Images 2.0 A state-of-the-art image model that can take on complex visual tasks and produce precise, immediately usable …

Model ReleasesDGX agent

Introducing ChatGPT Images 2.0 A state-of-the-art image model that can take on complex visual tasks and produce precise, immediately usable visuals, with sharper editing, richer layouts, and thinking-

Large Language Models Are Still Misled by Simple Bias Ensembles

Model ReleasesDGX agent

arXiv:2505.16522v3 Announce Type: replace Abstract: With the evolution of large language models (LLMs), their robustness against individual simple biases has been enhanced. However, we observe that th

Learning to Seek Help: Dynamic Collaboration Between Small and Large Language Models

Local AiDGX agent

arXiv:2604.17827v1 Announce Type: new Abstract: Large language models (LLMs) offer strong capabilities but raise cost and privacy concerns, whereas small language models (SLMs) facilitate efficient an

Multilingual Training and Evaluation Resources for Vision-Language Models

ResearchDGX agent

arXiv:2604.18347v1 Announce Type: new Abstract: Vision Language Models (VLMs) achieved rapid progress in the recent years. However, despite their growth, VLMs development is heavily grounded on Englis

Non-Stationarity in the Embedding Space of Time Series Foundation Models

ResearchDGX agent

arXiv:2604.16428v1 Announce Type: new Abstract: Time series foundation models (TSFMs) are widely used as generic feature extractors, yet the notion of non-stationarity in their embedding spaces remain

PDDL-Mind: Large Language Models are Capable on Belief Reasoning with Reliable State Tracking

Model ReleasesDGX agent

arXiv:2604.17819v1 Announce Type: new Abstract: Large language models (LLMs) perform substantially below human level on existing theory-of-mind (ToM) benchmarks, even when augmented with chain-of-thou

Prior-Fitted Functional Flow: In-Context Generative Models for Pharmacokinetics

Model ReleasesDGX agent

arXiv:2604.17670v1 Announce Type: new Abstract: We introduce Prior-Fitted Functional Flows, a generative foundation model for pharmacokinetics that enables zero-shot population synthesis and individua

Reducing Peak Memory Usage for Modern Multimodal Large Language Model Pipelines

ResearchDGX agent

arXiv:2604.16734v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have recently demonstrated strong capabilities in understanding and generating responses from diverse visual in

REFLEX: Reference-Free Evaluation of Log Summarization via Large Language Model Judgment

ApplicationsDGX agent

arXiv:2511.07458v2 Announce Type: replace Abstract: Evaluating log summarization systems is challenging due to the lack of high-quality reference summaries and the limitations of existing metrics like

Representation Before Training: A Fixed-Budget Benchmark for Generative Medical Event Models

Model ReleasesDGX agent

arXiv:2604.16775v1 Announce Type: new Abstract: Every prediction from a generative medical event model is bounded by how clinical events are tokenized, yet input representation is rarely isolated from

Retrieval-Augmented Multimodal Model for Fake News Detection

SafetyDGX agent

arXiv:2604.18112v1 Announce Type: new Abstract: In recent years, multimodal multidomain fake news detection has garnered increasing attention. Nevertheless, this direction presents two significant cha

Revisiting a Pain in the Neck: A Semantic Reasoning Benchmark for Language Models

Model ReleasesDGX agent

arXiv:2604.16593v1 Announce Type: new Abstract: We present SemanticQA, an evaluation suite designed to assess language models (LMs) in semantic phrase processing tasks. The benchmark consolidates exis

SHRUG-FM: Reliability-Aware Foundation Models for Earth Observation

Model ReleasesDGX agent

arXiv:2511.10370v2 Announce Type: replace Abstract: Geospatial foundation models (GFMs) for Earth observation often fail to perform reliably in environments underrepresented during pretraining. We int

SpeechMedAssist: Efficiently and Effectively Adapting Speech Language Models for Medical Consultation

Model ReleasesDGX agent

arXiv:2601.04638v2 Announce Type: replace Abstract: Medical consultations are intrinsically speech-centric. However, most prior works focus on long-text-based interactions, which are cumbersome and pa

SPS: Steering Probability Squeezing for Better Exploration in Reinforcement Learning for Large Language Models

SafetyDGX agent

arXiv:2604.16995v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a promising paradigm for training reasoning-oriented models by leveraging rule-based reward signals. However,

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement

Model ReleasesDGX agent

arXiv:2604.17887v1 Announce Type: new Abstract: Inverse Dynamics Models (IDMs) map visual observations to low-level action commands, serving as central components for data labeling and policy executio

StableMTL: Repurposing Latent Diffusion Models for Multi-Task Learning from Partially Annotated Synthetic Datasets

ResearchDGX agent

arXiv:2506.08013v2 Announce Type: replace Abstract: Multi-task learning for dense prediction is limited by the need for extensive annotation for every task, though recent works have explored training

StrEBM: A Structured Latent Energy-Based Model for Blind Source Separation

ResearchDGX agent

arXiv:2604.17381v1 Announce Type: cross Abstract: This paper proposes StrEBM, a structured latent energy-based model for source-wise structured representation learning. The framework is motivated by a

SynopticBench: Evaluating Vision-Language Models on Generating Weather Forecast Discussions of the Future

SafetyDGX agent

arXiv:2604.16451v1 Announce Type: new Abstract: Recent advances in visual-language models (VLMs) have led to significant improvements in a plethora of complex multimodal tasks like image captioning, r

Test-Time Perturbation Learning with Delayed Feedback for Vision-Language-Action Models

ResearchDGX agent

arXiv:2604.18107v1 Announce Type: new Abstract: Vision-Language-Action models (VLAs) achieve remarkable performance in sequential decision-making but remain fragile to subtle environmental shifts, suc

The Global Neural World Model: Spatially Grounded Discrete Topologies for Action-Conditioned Planning

AgentsDGX agent

arXiv:2604.16585v1 Announce Type: new Abstract: We present the Global Neural World Model (GNWM), a self-stabilizing framework that achieves topological quantization through balanced continuous entropy

TLoRA: Task-aware Low Rank Adaptation of Large Language Models

Model ReleasesDGX agent

arXiv:2604.18124v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become a widely adopted parameter-efficient fine-tuning method for large language models, with its effectiveness largely

Towards Joint Quantization and Token Pruning of Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.17320v1 Announce Type: new Abstract: Deploying Vision-Language Models (VLMs) under aggressive low-bit inference remains challenging because inference cost is dominated by the long visual-to

Unsupervised Discovery of Intermediate Phase Order in the Frustrated J_1-J_2 Heisenberg Model via Prometheus Framework

Model ReleasesDGX agent

arXiv:2602.21468v4 Announce Type: replace-cross Abstract: The spin-1/2 J_1-J_2 Heisenberg model on the square lattice exhibits a debated intermediate phase between Neel antiferromagnetic and stripe or

When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models

Model ReleasesDGX agent

arXiv:2604.17375v1 Announce Type: new Abstract: Recent advances in Vision-Language Models (VLMs) have substantially enhanced their ability across multimodal video understanding benchmarks spanning tem

20 Apr 2026

AI reviewers then ranked the submissions, and gave the same ordering every time, regardless of model doing the ranking: Codex GPT-5.4 > GPT-…

Model ReleasesDGX agent

AI reviewers then ranked the submissions, and gave the same ordering every time, regardless of model doing the ranking: Codex GPT-5.4 > GPT-5.3-Codex > Opus 4.6 > humans. Paper: http://claude-code-eco

Applied Explainability for Large Language Models: A Comparative Study

ApplicationsDGX agent

arXiv:2604.15371v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong performance across many natural language processing tasks, yet their decision processes remain difficult t

AutoDrive-R^2: Incentivizing Reasoning and Self-Reflection Capacity for VLA Model in Autonomous Driving

SafetyDGX agent

arXiv:2509.01944v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models in autonomous driving systems have recently demonstrated transformative potential by integrating multimoda

Automating Crash Diagram Generation Using Vision-Language Models: A Case Study on Multi-Lane Roundabouts

Model ReleasesDGX agent

arXiv:2604.15332v1 Announce Type: cross Abstract: Crash diagrams are essential tools in transportation safety analysis, yet their manual preparation remains time-consuming and prone to human variabili

Concept-wise Attention for Fine-grained Concept Bottleneck Models

SafetyDGX agent

arXiv:2604.15748v1 Announce Type: new Abstract: Recently impressive performance has been achieved in Concept Bottleneck Models (CBM) by utilizing the image-text alignment learned by a large pre-traine

ConFu: Contemplate the Future for Better Speculative Sampling

Model ReleasesDGX agent

arXiv:2603.08899v2 Announce Type: replace Abstract: Speculative decoding has emerged as a powerful approach to accelerate large language model (LLM) inference by employing lightweight draft models to

DINOv3 Beats Specialized Detectors: A Simple Foundation Model Baseline for Image Forensics

Local AiDGX agent

arXiv:2604.16083v1 Announce Type: new Abstract: With the rapid advancement of deep generative models, realistic fake images have become increasingly accessible, yet existing localization methods rely

FETAL-GAUGE: A Benchmark for Assessing Vision-Language Models in Fetal Ultrasound

Model ReleasesDGX agent

arXiv:2512.22278v2 Announce Type: replace Abstract: The growing demand for prenatal ultrasound imaging has intensified a global shortage of trained sonographers, creating barriers to essential fetal h

Free ~20-50% tok/s on a local llama.cpp setup if you already have a draft model sharing vocabulary with your main one. Local stack quietly g…

Model ReleasesDGX agent

Free ~20-50% tok/s on a local llama.cpp setup if you already have a draft model sharing vocabulary with your main one. Local stack quietly got faster this weekend https://x.com/TechIno219886/status/20

HyperGVL: Benchmarking and Improving Large Vision-Language Models in Hypergraph Understanding and Reasoning

Model ReleasesDGX agent

arXiv:2604.15648v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) consistently require new arenas to guide their expanding boundaries, yet their capabilities with hypergraphs remain

I upgraded my Claude token counter tool to compare different models and Opus 4.7 does appear to use 1.46x times the tokens for text and up t…

Model ReleasesDGX agent

I upgraded my Claude token counter tool to compare different models and Opus 4.7 does appear to use 1.46x times the tokens for text and up to 3x the tokens for images - it's priced the same as Opus 4.

Jailbreak Scaling Laws for Large Language Models: Polynomial-Exponential Crossover

SafetyDGX agent

arXiv:2603.11331v2 Announce Type: replace-cross Abstract: Adversarial attacks can reliably steer safety-aligned large language models toward unsafe behavior. Empirically, we find that strong adversari

JumpLoRA: Sparse Adapters for Continual Learning in Large Language Models

Model ReleasesDGX agent

arXiv:2604.16171v1 Announce Type: cross Abstract: Adapter-based methods have become a cost-effective approach to continual learning (CL) for Large Language Models (LLMs), by sequentially learning a lo

life when you discover an open-source model that runs 300 parallel agents, executes for 12+ hours straight, beats GPT-5.4 and opus 4.6 on mu…

Model ReleasesDGX agent

life when you discover an open-source model that runs 300 parallel agents, executes for 12+ hours straight, beats GPT-5.4 and opus 4.6 on multiple benchmarks... and the weights are on huggingface Medi

Reasoning-targeted Jailbreak Attacks on Large Reasoning Models via Semantic Triggers and Psychological Framing

Model ReleasesDGX agent

arXiv:2604.15725v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have demonstrated strong capabilities in generating step-by-step reasoning chains alongside final answers, enabling thei

← Previous
1…119120121122123…1009
Next →