AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,565 results
Model Releases

Once Correct, Still Wrong: Counterfactual Hallucination in Multilingual Vision-Language Models

DGX agent

arXiv:2602.05437v2 Announce Type: replace Abstract: Vision-language models (VLMs) can achieve high accuracy while still accepting culturally plausible but visually incorrect interpretations. Existing

model-releasesarxiv-cs-cl
22 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Applications

ParamBoost: Gradient Boosted Piecewise Cubic Polynomials

DGX agent

arXiv:2604.18864v1 Announce Type: new Abstract: Generalized Additive Models (GAMs) can be used to create non-linear glass-box (i.e. explicitly interpretable) models, where the predictive function is f

applicationsarxiv-cs-lg
22 Apr 2026
Model Releases

Presumably GPT-imagegen-2 (aka ChatGPT Images 2.0 aka gpt-image-2) works as a tool which the models generate prompts for? I wish we could se…

DGX agent

Presumably GPT-imagegen-2 (aka ChatGPT Images 2.0 aka gpt-image-2) works as a tool which the models generate prompts for? I wish we could see those prompts, like back in the DALL-E 3 days https://simo

model-releasessimon-willison--x
22 Apr 2026
Safety

ProjLens: Unveiling the Role of Projectors in Multimodal Model Safety

DGX agent

arXiv:2604.19083v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success in cross-modal understanding and generation, yet their deployment is threate

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Q-Mask: Query-driven Causal Masks for Text Anchoring in OCR-Oriented Vision-Language Models

DGX agent

arXiv:2604.00161v2 Announce Type: replace Abstract: Optical Character Recognition (OCR) is increasingly regarded as a foundational capability for modern vision-language models (VLMs), enabling them no

model-releasesarxiv-cs-cv
22 Apr 2026
Local Ai

R^2-dLLM: Accelerating Diffusion Large Language Models via Spatio-Temporal Redundancy Reduction

DGX agent

arXiv:2604.18995v1 Announce Type: cross Abstract: Diffusion Large Language Models (dLLMs) have emerged as a promising alternative to autoregressive generation by enabling parallel token prediction. Ho

local-aiarxiv-cs-ai
22 Apr 2026
Agents

Rethinking Scale: Deployment Trade-offs of Small Language Models under Agent Paradigms

DGX agent

arXiv:2604.19299v1 Announce Type: cross Abstract: Despite the impressive capabilities of large language models, their substantial computational costs, latency, and privacy risks hinder their widesprea

agentsarxiv-cs-ai
22 Apr 2026
Safety

Safety-Critical Contextual Control via Online Riemannian Optimization with World Models

DGX agent

arXiv:2604.19639v1 Announce Type: cross Abstract: Modern world models are becoming too complex to admit explicit dynamical descriptions. We study safety-critical contextual control, where a Planner mu

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Sources: OpenAI has been briefing US federal agencies, state governments, and Five Eyes allies on the capabilities of its GPT-5.4-Cyber model over the past week (Sam Sabin/Axios)

DGX agent

Sam Sabin / Axios: Sources: OpenAI has been briefing US federal agencies, state governments, and Five Eyes allies on the capabilities of its GPT-5.4-Cyber model over the past week — OpenAI has been br

model-releasestechmeme
22 Apr 2026
Safety

TEMPO: Scaling Test-time Training for Large Reasoning Models

DGX agent

arXiv:2604.19295v1 Announce Type: new Abstract: Test-time training (TTT) adapts model parameters on unlabeled test instances during inference time, which continuously extends capabilities beyond the r

safetyarxiv-cs-lg
22 Apr 2026
Model Releases

What’s new in Cloud Run at Next ‘26

DGX agent

From vibe-coded and large-scale apps to AI models and agents, Cloud Run delivers on-demand compute with zero overhead and pay-per-use pricing for all of your workloads. Last year, the number of extern

model-releasesgoogle-cloud-ai
22 Apr 2026
Model Releases

A Systematic Survey and Benchmark of Deep Learning for Molecular Property Prediction in the Foundation Model Era

DGX agent

arXiv:2604.16586v1 Announce Type: new Abstract: Molecular property prediction integrates quantum chemistry, cheminformatics, and deep learning to connect molecular structure with physicochemical and b

model-releasesarxiv-cs-lg
21 Apr 2026
Research

A Unification of Discrete, Gaussian, and Simplicial Diffusion

DGX agent

arXiv:2512.15923v2 Announce Type: replace Abstract: To model discrete sequences such as DNA, proteins, and language using diffusion, practitioners must choose between three major methods: diffusion in

researcharxiv-cs-lg
21 Apr 2026
Applications

An LLM-Guided Query-Aware Inference System for GNN Models on Large Knowledge Graphs

DGX agent

arXiv:2603.04545v2 Announce Type: replace Abstract: Efficient inference for graph neural networks (GNNs) on large knowledge graphs (KGs) is essential for many real-world applications. GNN inference qu

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

BEFT: Bias-Efficient Fine-Tuning of Language Models in Low-Data Regimes

DGX agent

arXiv:2509.15974v2 Announce Type: replace Abstract: Fine-tuning the bias terms of large language models (LLMs) has the potential to achieve unprecedented parameter efficiency while maintaining competi

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Calibrating Model-Based Evaluation Metrics for Summarization

DGX agent

arXiv:2604.17200v1 Announce Type: new Abstract: Recent advances in summary evaluation are based on model-based metrics to assess quality dimensions, such as completeness, conciseness, and faithfulness

researcharxiv-cs-cl
21 Apr 2026
Safety

Closing the Modality Reasoning Gap for Speech Large Language Models

DGX agent

arXiv:2601.05543v2 Announce Type: replace Abstract: Although Speech Large Language Models have achieved notable progress, a substantial modality reasoning gap remains: their reasoning performance on s

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

ConMeZO: Adaptive Descent-Direction Sampling for Gradient-Free Finetuning of Large Language Models

DGX agent

arXiv:2511.02757v2 Announce Type: replace Abstract: Zeroth-order or derivative-free optimization (MeZO) is an attractive strategy for finetuning large language models (LLMs) because it eliminates the

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Domain-Specialized Object Detection via Model-Level Mixtures of Experts

DGX agent

arXiv:2604.18256v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models provide a structured approach to combining specialized neural networks and offer greater interpretability than conventio

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Employing General-Purpose and Biomedical Large Language Models with Advanced Prompt Engineering for Pharmacoepidemiologic Study Design

DGX agent

arXiv:2604.17988v1 Announce Type: new Abstract: Background: The potential of large language models (LLMs) to automate and support pharmacoepidemiologic study design is an emerging area of interest, ye

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Erasing Thousands of Concepts: Towards Scalable and Practical Concept Erasure for Text-to-Image Diffusion Models

DGX agent

arXiv:2604.16481v1 Announce Type: new Abstract: Large-scale text-to-image (T2I) diffusion models deliver remarkable visual fidelity but pose safety risks due to their capacity to reproduce undesirable

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection

DGX agent

arXiv:2410.04509v3 Announce Type: replace Abstract: As the field of Multimodal Large Language Models (MLLMs) continues to evolve, their potential to revolutionize artificial intelligence is particular

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Establishing a Scale for Kullback-Leibler Divergence in Language Models Across Various Settings

DGX agent

arXiv:2505.15353v3 Announce Type: replace Abstract: Log-likelihood vectors define a common space for comparing language models as probability distributions, enabling unified comparisons across heterog

researcharxiv-cs-cl
21 Apr 2026
Model Releases

FedLLM: A Privacy-Preserving Federated Large Language Model for Explainable Traffic Flow Prediction

DGX agent

arXiv:2604.16612v1 Announce Type: new Abstract: Traffic prediction plays a central role in intelligent transportation systems (ITS) by supporting real-time decision-making, congestion management, and

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

FLARE: Task-agnostic embedding model evaluation through a normalization process

DGX agent

arXiv:2604.17344v1 Announce Type: cross Abstract: When task-specific labels are not available, it becomes difficult to select an embedding model for a specific target corpus. Existing labelless measur

model-releasesarxiv-cs-cl
21 Apr 2026
Research

GaLa: Hypergraph-Guided Visual Language Models for Procedural Planning

DGX agent

arXiv:2604.17241v1 Announce Type: new Abstract: Implicit spatial relations and deep semantic structures encoded in object attributes are crucial for procedural planning in embodied AI systems. However

researcharxiv-cs-ro
21 Apr 2026
Tutorials

Graph neural network for colliding particles with an application to sea ice floe modeling

DGX agent

arXiv:2602.16213v2 Announce Type: replace-cross Abstract: This paper introduces a novel approach to sea ice modeling using Graph Neural Networks (GNNs), utilizing the natural graph structure of sea ic

tutorialsarxiv-cs-cv
21 Apr 2026
Safety

How Language Models Conflate Logical Validity with Plausibility: A Representational Analysis of Content Effects

DGX agent

arXiv:2510.06700v3 Announce Type: replace Abstract: Both humans and large language models (LLMs) exhibit content effects: biases in which the plausibility of the semantic content of a reasoning proble

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study

DGX agent

arXiv:2505.15404v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have achieved remarkable success on reasoning-intensive tasks such as mathematics and programming. However, their enha

model-releasesarxiv-cs-cl
21 Apr 2026
Tutorials

How to Approximate Inference with Subtractive Mixture Models

DGX agent

arXiv:2604.16714v1 Announce Type: new Abstract: Classical mixture models (MMs) are widely used tractable proposals for approximate inference settings such as variational inference (VI) and importance

tutorialsarxiv-cs-lg
21 Apr 2026
Model Releases

Introducing ChatGPT Images 2.0 A state-of-the-art image model that can take on complex visual tasks and produce precise, immediately usable …

DGX agent

Introducing ChatGPT Images 2.0 A state-of-the-art image model that can take on complex visual tasks and produce precise, immediately usable visuals, with sharper editing, richer layouts, and thinking-

model-releasesopenai--x
21 Apr 2026
Model Releases

Large Language Models Are Still Misled by Simple Bias Ensembles

DGX agent

arXiv:2505.16522v3 Announce Type: replace Abstract: With the evolution of large language models (LLMs), their robustness against individual simple biases has been enhanced. However, we observe that th

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

Learning to Seek Help: Dynamic Collaboration Between Small and Large Language Models

DGX agent

arXiv:2604.17827v1 Announce Type: new Abstract: Large language models (LLMs) offer strong capabilities but raise cost and privacy concerns, whereas small language models (SLMs) facilitate efficient an

local-aiarxiv-cs-cl
21 Apr 2026
Research

Multilingual Training and Evaluation Resources for Vision-Language Models

DGX agent

arXiv:2604.18347v1 Announce Type: new Abstract: Vision Language Models (VLMs) achieved rapid progress in the recent years. However, despite their growth, VLMs development is heavily grounded on Englis

researcharxiv-cs-cl
21 Apr 2026
Research

Non-Stationarity in the Embedding Space of Time Series Foundation Models

DGX agent

arXiv:2604.16428v1 Announce Type: new Abstract: Time series foundation models (TSFMs) are widely used as generic feature extractors, yet the notion of non-stationarity in their embedding spaces remain

researcharxiv-cs-lg
21 Apr 2026
Model Releases

PDDL-Mind: Large Language Models are Capable on Belief Reasoning with Reliable State Tracking

DGX agent

arXiv:2604.17819v1 Announce Type: new Abstract: Large language models (LLMs) perform substantially below human level on existing theory-of-mind (ToM) benchmarks, even when augmented with chain-of-thou

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Prior-Fitted Functional Flow: In-Context Generative Models for Pharmacokinetics

DGX agent

arXiv:2604.17670v1 Announce Type: new Abstract: We introduce Prior-Fitted Functional Flows, a generative foundation model for pharmacokinetics that enables zero-shot population synthesis and individua

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Reducing Peak Memory Usage for Modern Multimodal Large Language Model Pipelines

DGX agent

arXiv:2604.16734v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have recently demonstrated strong capabilities in understanding and generating responses from diverse visual in

researcharxiv-cs-cv
21 Apr 2026
Applications

REFLEX: Reference-Free Evaluation of Log Summarization via Large Language Model Judgment

DGX agent

arXiv:2511.07458v2 Announce Type: replace Abstract: Evaluating log summarization systems is challenging due to the lack of high-quality reference summaries and the limitations of existing metrics like

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

Representation Before Training: A Fixed-Budget Benchmark for Generative Medical Event Models

DGX agent

arXiv:2604.16775v1 Announce Type: new Abstract: Every prediction from a generative medical event model is bounded by how clinical events are tokenized, yet input representation is rarely isolated from

model-releasesarxiv-cs-lg
21 Apr 2026
Safety

Retrieval-Augmented Multimodal Model for Fake News Detection

DGX agent

arXiv:2604.18112v1 Announce Type: new Abstract: In recent years, multimodal multidomain fake news detection has garnered increasing attention. Nevertheless, this direction presents two significant cha

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Revisiting a Pain in the Neck: A Semantic Reasoning Benchmark for Language Models

DGX agent

arXiv:2604.16593v1 Announce Type: new Abstract: We present SemanticQA, an evaluation suite designed to assess language models (LMs) in semantic phrase processing tasks. The benchmark consolidates exis

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

SHRUG-FM: Reliability-Aware Foundation Models for Earth Observation

DGX agent

arXiv:2511.10370v2 Announce Type: replace Abstract: Geospatial foundation models (GFMs) for Earth observation often fail to perform reliably in environments underrepresented during pretraining. We int

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

SpeechMedAssist: Efficiently and Effectively Adapting Speech Language Models for Medical Consultation

DGX agent

arXiv:2601.04638v2 Announce Type: replace Abstract: Medical consultations are intrinsically speech-centric. However, most prior works focus on long-text-based interactions, which are cumbersome and pa

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

SPS: Steering Probability Squeezing for Better Exploration in Reinforcement Learning for Large Language Models

DGX agent

arXiv:2604.16995v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a promising paradigm for training reasoning-oriented models by leveraging rule-based reward signals. However,

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement

DGX agent

arXiv:2604.17887v1 Announce Type: new Abstract: Inverse Dynamics Models (IDMs) map visual observations to low-level action commands, serving as central components for data labeling and policy executio

model-releasesarxiv-cs-ro
21 Apr 2026
Research

StableMTL: Repurposing Latent Diffusion Models for Multi-Task Learning from Partially Annotated Synthetic Datasets

DGX agent

arXiv:2506.08013v2 Announce Type: replace Abstract: Multi-task learning for dense prediction is limited by the need for extensive annotation for every task, though recent works have explored training

researcharxiv-cs-cv
21 Apr 2026
Research

StrEBM: A Structured Latent Energy-Based Model for Blind Source Separation

DGX agent

arXiv:2604.17381v1 Announce Type: cross Abstract: This paper proposes StrEBM, a structured latent energy-based model for source-wise structured representation learning. The framework is motivated by a

researcharxiv-cs-lg
21 Apr 2026
← Previous
1…149150151152153…1262
Next →