AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlog
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,490 results
Research

My bets on open models, mid-2026

DGX agent

A mid-2026 outlook piece from the Interconnects AI newsletter in which the author makes specific predictions about the trajectory of open-weight language models, likely covering expected capability mi

researchinterconnects
15 Apr 2026
Model Releases

OFA-Diffusion Compression: Compressing Diffusion Model in One-Shot Manner

DGX agent
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

arXiv:2604.12668v1 Announce Type: new Abstract: The Diffusion Probabilistic Model (DPM) achieves remarkable performance in image generation, while its increasing parameter size and computational overh

model-releasesarxiv-cs-cv
15 Apr 2026
Research

Scaling Exposes the Trigger: Input-Level Backdoor Detection in Text-to-Image Diffusion Models via Cross-Attention Scaling

DGX agent

arXiv:2604.12446v1 Announce Type: cross Abstract: Text-to-image (T2I) diffusion models have achieved remarkable success in image synthesis, but their reliance on large-scale data and open ecosystems i

researcharxiv-cs-cv
15 Apr 2026
Model Releases

SinkSAM-Net: Knowledge-Driven Self-Supervised Sinkhole Segmentation Using Topographic Priors and Segment Anything Model

DGX agent

arXiv:2410.01473v2 Announce Type: replace Abstract: Soil sinkholes significantly influence soil degradation, infrastructure vulnerability, and landscape evolution. However, their irregular shapes, com

model-releasesarxiv-cs-cv
15 Apr 2026
Agents

Small models are cheap to run, but expensive to adapt. The hard part is not only fine-tuning. It is the surrounding loop that involves colle…

DGX agent

Small models are cheap to run, but expensive to adapt. The hard part is not only fine-tuning. It is the surrounding loop that involves collecting data, diagnosing failures, building evals, avoiding re

agentsdair-ai--x
15 Apr 2026
Safety

SOAR: Self-Correction for Optimal Alignment and Refinement in Diffusion Models

DGX agent

arXiv:2604.12617v1 Announce Type: cross Abstract: The post-training pipeline for diffusion models currently has two stages: supervised fine-tuning (SFT) on curated data and reinforcement learning (RL)

safetyarxiv-cs-ai
15 Apr 2026
Local Ai

Towards Interpretable Foundation Models for Retinal Fundus Images

DGX agent

arXiv:2603.18846v2 Announce Type: replace Abstract: Foundation models are used to extract transferable representations from large amounts of unlabeled data, typically via self-supervised learning (SSL

local-aiarxiv-cs-cv
15 Apr 2026
Local Ai

What does 'Run <number> cloud models at a time' in Ollama Cloud Subscription mean?

DGX agent

The Ollama Cloud Subscription includes a feature described as 'Run cloud models at a time,' which refers to how many AI models a user can have simultaneously loaded and running in the cloud at any giv

local-air-ollama
15 Apr 2026
Model Releases

A Compact and Efficient 1.251 Million Parameter Machine Learning CNN Model PD36-C for Plant Disease Detection: A Case Study

DGX agent

arXiv:2604.11332v1 Announce Type: cross Abstract: Deep learning has markedly advanced image based plant disease diagnosis as improved hardware and dataset quality have enabled increasingly accurate ne

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Anthropic details using AI agents to accelerate alignment research on 'weak-to-strong supervision', where a weak model supervises the training of a stronger one (Anthropic)

DGX agent

Anthropic: Anthropic details using AI agents to accelerate alignment research on “weak-to-strong supervision”, where a weak model supervises the training of a stronger one — Large language models' eve

safetytechmeme
14 Apr 2026
Safety

Anthropogenic Regional Adaptation in Multimodal Vision-Language Model

DGX agent

arXiv:2604.11490v1 Announce Type: new Abstract: While the field of vision-language (VL) has achieved remarkable success in integrating visual and textual information across multiple languages and doma

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

AOP-Smart: A RAG-Enhanced Large Language Model Framework for Adverse Outcome Pathway Analysis

DGX agent

arXiv:2604.10874v1 Announce Type: cross Abstract: Adverse Outcome Pathways (AOPs) are an important knowledge framework in toxicological research and risk assessment. In recent years, large language mo

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Benchmarking Large Vision-Language Models on Fine-Grained Image Tasks: A Comprehensive Evaluation

DGX agent

arXiv:2504.14988v3 Announce Type: replace Abstract: Recent advancements in Large Vision-Language Models (LVLMs) have demonstrated remarkable multimodal perception capabilities, garnering significant a

model-releasesarxiv-cs-cv
14 Apr 2026
Research

COREY: A Prototype Study of Entropy-Guided Operator Fusion with Hadamard Reparameterization for Selective State Space Models

DGX agent

arXiv:2604.10597v1 Announce Type: cross Abstract: State Space Models (SSMs), represented by the Mamba family, provide linear-time sequence modeling and are attractive for long-context inference. Yet p

researcharxiv-cs-ai
14 Apr 2026
Research

Data-Efficient Surgical Phase Segmentation in Small-Incision Cataract Surgery: A Controlled Study of Vision Foundation Models

DGX agent

arXiv:2604.10514v1 Announce Type: cross Abstract: Surgical phase segmentation is central to computer-assisted surgery, yet robust models remain difficult to develop when labeled surgical videos are sc

researcharxiv-cs-ai
14 Apr 2026
Safety

Demographic and Linguistic Bias Evaluation in Omnimodal Language Models

DGX agent

arXiv:2604.10014v1 Announce Type: cross Abstract: This paper provides a comprehensive evaluation of demographic and linguistic biases in omnimodal language models that process text, images, audio, and

safetyarxiv-cs-ai
14 Apr 2026
Safety

Early Decisions Matter: Proximity Bias and Initial Trajectory Shaping in Non-Autoregressive Diffusion Language Models

DGX agent

arXiv:2604.10567v1 Announce Type: cross Abstract: Diffusion-based language models (dLLMs) have emerged as a promising alternative to autoregressive language models, offering the potential for parallel

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Finetune Like You Pretrain: Boosting Zero-shot Adversarial Robustness in Vision-language Models

DGX agent

arXiv:2604.11576v1 Announce Type: new Abstract: Despite their impressive zero-shot abilities, vision-language models such as CLIP have been shown to be susceptible to adversarial attacks. To enhance i

model-releasesarxiv-cs-cv
14 Apr 2026
Agents

For an agent builder, memory is sustained advantage For the model provider, memory is switching cost

DGX agent

Harrison Chase's post explores the divergent strategic incentives around memory in AI agent systems, arguing that for those building agent frameworks, persistent memory creates compounding value and c

agentsharrison-chase--x
14 Apr 2026
Model Releases

FPBench: A Comprehensive Benchmark of Multimodal Large Language Models for Fingerprint Analysis

DGX agent

arXiv:2512.18073v2 Announce Type: replace Abstract: Multimodal LLMs (MLLMs) are capable of performing complex data analysis, visual question answering, generation, and reasoning tasks. However, their

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

From Topology to Trajectory: LLM-Driven World Models For Supply Chain Resilience

DGX agent

arXiv:2604.11041v1 Announce Type: new Abstract: Semiconductor supply chains face unprecedented resilience challenges amidst global geopolitical turbulence. Conventional Large Language Model (LLM) plan

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

GeoMeld: Toward Semantically Grounded Foundation Models for Remote Sensing

DGX agent

arXiv:2604.10591v1 Announce Type: cross Abstract: Effective foundation modeling in remote sensing requires spatially aligned heterogeneous modalities coupled with semantically grounded supervision, ye

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

LoGo-MR: Screening Breast MRI for Cancer Risk Prediction by Efficient Omni-Slice Modeling

DGX agent

arXiv:2604.11348v1 Announce Type: new Abstract: Efficient and explainable breast cancer (BC) risk prediction is critical for large-scale population-based screening. Breast MRI provides functional info

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Measuring the Authority Stack of AI Systems: Empirical Analysis of 366,120 Forced-Choice Responses Across 8 AI Models

DGX agent

arXiv:2604.11216v1 Announce Type: new Abstract: What values, evidence preferences, and source trust hierarchies do AI systems actually exhibit when facing structured dilemmas? We present the first lar

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

MEDSYN: Benchmarking Multi-EviDence SYNthesis in Complex Clinical Cases for Multimodal Large Language Models

DGX agent

arXiv:2602.21950v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have shown great potential in medical applications, yet existing benchmarks inadequately capture real-world

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

METER: Evaluating Multi-Level Contextual Causal Reasoning in Large Language Models

DGX agent

arXiv:2604.11502v1 Announce Type: cross Abstract: Contextual causal reasoning is a critical yet challenging capability for Large Language Models (LLMs). Existing benchmarks, however, often evaluate th

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

MMR-AD: A Large-Scale Multimodal Dataset for Benchmarking General Anomaly Detection with Multimodal Large Language Models

DGX agent

arXiv:2604.10971v1 Announce Type: cross Abstract: In the progress of industrial anomaly detection, general anomaly detection (GAD) is an emerging trend and also the ultimate goal. Unlike the conventio

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

NEW: Added @OpenRouter's new free stealth model, Elephant-Alpha, to Hermes Agent, which you can now access if you run `hermes update`! I als…

DGX agent

NEW: Added @OpenRouter's new free stealth model, Elephant-Alpha, to Hermes Agent, which you can now access if you run `hermes update`! I also had Hermes come up with an agentic benchmark on the fly to

model-releasesnous-research--x
14 Apr 2026
Hardware

Nvidia unveils Ising AI models for quantum error correction and calibration

DGX agent

Technology and computing giant Nvidia Corp. today announced the release of Ising, the world’s first open artificial intelligence model family aimed at quantum computing calibration and error correctio

hardwaresiliconangle
14 Apr 2026
Model Releases

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models

DGX agent

arXiv:2604.10866v1 Announce Type: new Abstract: AI agents are expected to perform professional work across hundreds of occupational domains (from emergency department triage to nuclear reactor safety

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

OpenAI launches GPT-5.4-Cyber model for vetted security pros

DGX agent

OpenAI Group PBC today announced the launch of GPT-5.4-Cyber, a fine-tuned variant of its GPT-5.4 model designed for defensive cybersecurity work and also announced a significant expansion of its Trus

model-releasessiliconangle
14 Apr 2026
Research

Physics and causally constrained discrete-time neural models of turbulent dynamical systems

DGX agent

arXiv:2602.13847v3 Announce Type: replace-cross Abstract: We present a framework for constructing physics and causally constrained neural models of turbulent dynamical systems from data. We first form

researcharxiv-cs-lg
14 Apr 2026
Research

PoreDiT: A Scalable Generative Model for Large-Scale Digital Rock Reconstruction

DGX agent

arXiv:2604.10171v1 Announce Type: new Abstract: This manuscript presents PoreDiT, a novel generative model designed for high-efficiency digital rock reconstruction at gigavoxel scales. Addressing the

researcharxiv-cs-ai
14 Apr 2026
Research

Resource Consumption Threats in Large Language Models

DGX agent

arXiv:2603.16068v3 Announce Type: replace-cross Abstract: Given limited and costly computational infrastructure, resource efficiency is a key requirement for large language models (LLMs). Efficient LL

researcharxiv-cs-ai
14 Apr 2026
Research

Rethinking the Diffusion Model from a Langevin Perspective

DGX agent

arXiv:2604.10465v1 Announce Type: cross Abstract: Diffusion models are often introduced from multiple perspectives, such as VAEs, score matching, or flow matching, accompanied by dense and technically

researcharxiv-cs-ai
14 Apr 2026
Model Releases

SimBench: Benchmarking the Ability of Large Language Models to Simulate Human Behaviors

DGX agent

arXiv:2510.17516v4 Announce Type: replace-cross Abstract: Large language model (LLM) simulations of human behavior have the potential to revolutionize the social and behavioral sciences, if and only i

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

The future is open. At GTC Live, @cohere Co-founder & CEO, Aidan Gomez describes how open models give developers the autonomy they need to s…

DGX agent

The future is open. At GTC Live, @cohere Co-founder & CEO, Aidan Gomez describes how open models give developers the autonomy they need to scale 📈, customize and innovate their way into the future of

agentscohere--x
14 Apr 2026
Agents

Today we're releasing SWE-check, a specialized bug detection model we RL-trained with @appliedcompute that matches frontier performance on i…

DGX agent

Today we're releasing SWE-check, a specialized bug detection model we RL-trained with @appliedcompute that matches frontier performance on internal in-distribution evals and makes meaningful progress

agentscognition-ai--x
14 Apr 2026
Model Releases

Too Nice to Tell the Truth: Quantifying Agreeableness-Driven Sycophancy in Role-Playing Language Models

DGX agent

arXiv:2604.10733v1 Announce Type: cross Abstract: Large language models increasingly serve as conversational agents that adopt personas and role-play characters at user request. This capability, while

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

TraversalBench: Challenging Paths to Follow for Vision Language Models

DGX agent

arXiv:2604.10999v1 Announce Type: new Abstract: Vision-language models (VLMs) perform strongly on many multimodal benchmarks. However, the ability to follow complex visual paths -- a task that human o

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Vision-Language-Action Model, Robustness, Multi-modal Learning, Robot Manipulation

DGX agent

arXiv:2604.10055v1 Announce Type: new Abstract: Despite their strong performance in embodied tasks, recent Vision-Language-Action (VLA) models remain highly fragile under multimodal perturbations, whe

model-releasesarxiv-cs-ro
14 Apr 2026
Research

You can decompose models into a graph database [N]

DGX agent

This Reddit post from r/MachineLearning discusses the concept of decomposing machine learning models into a graph database representation, treating a model's components — such as layers, weights, and

researchr-machinelearning
14 Apr 2026
Model Releases

Your Model Diversity, Not Method, Determines Reasoning Strategy

DGX agent

arXiv:2604.10827v1 Announce Type: new Abstract: Compute scaling for LLM reasoning requires allocating budget between exploring solution approaches (breadth) and refining promising solutions (depth). M

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Arbitration Failure, Not Perceptual Blindness: How Vision-Language Models Resolve Visual-Linguistic Conflicts

DGX agent

arXiv:2604.09364v1 Announce Type: cross Abstract: When a Vision-Language Model (VLM) sees a blue banana and answers 'yellow', is the problem of perception or arbitration? We explore the question in te

researcharxiv-cs-cl
13 Apr 2026
Model Releases

CONDESION-BENCH: Conditional Decision-Making of Large Language Models in Compositional Action Space

DGX agent

arXiv:2604.09029v1 Announce Type: cross Abstract: Large language models have been widely explored as decision-support tools in high-stakes domains due to their contextual understanding and reasoning c

model-releasesarxiv-cs-ai
13 Apr 2026
Safety

From Dispersion to Attraction: Spectral Dynamics of Hallucination Across Whisper Model Scales

DGX agent

arXiv:2604.08591v1 Announce Type: cross Abstract: Hallucinations in large ASR models present a critical safety risk. In this work, we propose the extit{Spectral Sensitivity Theorem}, which predicts a

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

From Reasoning to Agentic: Credit Assignment in Reinforcement Learning for Large Language Models

DGX agent

arXiv:2604.09459v1 Announce Type: new Abstract: Reinforcement learning (RL) for large language models (LLMs) increasingly relies on sparse, outcome-level rewards -- yet determining which actions withi

model-releasesarxiv-cs-cl
13 Apr 2026
Local Ai

I’m looking for advice on setting up a local AI model that can generate Word reports automatically.

DGX agent

This r/ollama thread discusses community advice on configuring a locally-run AI model (via Ollama) to automatically generate Word documents or reports, covering topics such as model selection, scripti

local-air-ollama
13 Apr 2026
← Previous
1…9394959697…1261
Next →