AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,522 results
15 Apr 2026

StructDiff: A Structure-Preserving and Spatially Controllable Diffusion Model for Single-Image Generation

TutorialsDGX agent

arXiv:2604.12575v1 Announce Type: new Abstract: This paper introduces StructDiff, a generative framework based on a single-scale diffusion model for single-image generation. Single-image generation ai

Surrogate models for diffusion on graphs via sparse polynomials

ApplicationsDGX agent

arXiv:2502.06595v3 Announce Type: replace-cross Abstract: Diffusion kernels over graphs have been widely utilized as effective tools in various applications due to their ability to accurately model th

The Illusion of Fit: Spatially Resolved Assessment of Constitutive Model Validity in Elastography and Physics-Based Inverse Problems

Local AiDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2502.07415v2 Announce Type: replace-cross Abstract: Inferring the mechanical properties of soft tissues from measured deformations is a fundamental challenge in elastography. A rarely examined a

Toward Efficient and Robust Behavior Models for Multi-Agent Driving Simulation

Local AiDGX agent

arXiv:2512.05812v5 Announce Type: replace-cross Abstract: Scalable multi-agent driving simulation requires behavior models that are both realistic and computationally efficient. We address this by opt

14 Apr 2026

Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis

Model ReleasesDGX agent

arXiv:2604.10233v1 Announce Type: cross Abstract: 3D medical image analysis is of great importance in disease diagnosis and treatment. Recently, multimodal large language models (MLLMs) have exhibited

Agentic Exploration of PDE Spaces using Latent Foundation Models for Parameterized Simulations

Model ReleasesDGX agent

arXiv:2604.09584v1 Announce Type: new Abstract: Flow physics and more broadly physical phenomena governed by partial differential equations (PDEs), are inherently continuous, high-dimensional and ofte

AI keeps getting better but the last time the shape of the jagged frontier changed radically was o1 & the Reasoner A good mental model of th…

ApplicationsDGX agent

AI keeps getting better but the last time the shape of the jagged frontier changed radically was o1 & the Reasoner A good mental model of the coming months is that models get very good at the things t

[AINews] Top Local Models List - April 2026

ToolsDGX agent

a quiet day lets us check in on the local models scene

Assessing Privacy Preservation and Utility in Online Vision-Language Models

ResearchDGX agent

arXiv:2604.09695v1 Announce Type: cross Abstract: The increasing use of Online Vision Language Models (OVLMs) for processing images has introduced significant privacy risks, as individuals frequently

CausalGaze: Unveiling Hallucinations via Counterfactual Graph Intervention in Large Language Models

ResearchDGX agent

arXiv:2604.11087v1 Announce Type: new Abstract: Despite the groundbreaking advancements made by large language models (LLMs), hallucination remains a critical bottleneck for their deployment in high-s

ComSim: Building Scalable Real-World Robot Data Generation via Compositional Simulation

SafetyDGX agent

arXiv:2604.11386v1 Announce Type: cross Abstract: Recent advancements in foundational models, such as large language models and world models, have greatly enhanced the capabilities of robotics, enabli

Detection Is Cheap, Routing Is Learned: Why Refusal-Based Alignment Evaluation Fails

SafetyDGX agent

arXiv:2603.18280v2 Announce Type: replace-cross Abstract: Current alignment evaluation mostly measures whether models encode dangerous concepts and whether they refuse harmful requests. Both miss the

EdgeCIM: A Hardware-Software Co-Design for CIM-Based Acceleration of Small Language Models

Model ReleasesDGX agent

arXiv:2604.11512v1 Announce Type: cross Abstract: The growing demand for deploying Small Language Models (SLMs) on edge devices, including laptops, smartphones, and embedded platforms, has exposed fun

EditCrafter: Tuning-free High-Resolution Image Editing via Pretrained Diffusion Model

ResearchDGX agent

arXiv:2604.10268v1 Announce Type: new Abstract: We propose EditCrafter, a high-resolution image editing method that operates without tuning, leveraging pretrained text-to-image (T2I) diffusion models

Evaluating Small Open LLMs for Medical Question Answering: A Practical Framework

Model ReleasesDGX agent

arXiv:2604.10535v1 Announce Type: cross Abstract: Incorporating large language models (LLMs) in medical question answering demands more than high average accuracy: a model that returns substantively d

From Perception to Autonomous Computational Modeling: A Multi-Agent Approach

Model ReleasesDGX agent

arXiv:2604.06788v2 Announce Type: replace-cross Abstract: We present a solver-agnostic framework in which coordinated large language model (LLM) agents autonomously execute the complete computational

GenTac: Generative Modeling and Forecasting of Soccer Tactics

Model ReleasesDGX agent

arXiv:2604.11786v1 Announce Type: new Abstract: Modeling open-play soccer tactics is a formidable challenge due to the stochastic, multi-agent nature of the game. Existing computational approaches typ

Helical raises $10M to bridge the gap between foundation models and drug discovery decisions

IndustryDGX agent

Pharma artificial intelligence startup Helical Ltd. announced today that it has raised 10 million in new funding to expand its virtual AI lab platform, which turns biological foundation models into re

How Controllable Are Large Language Models? A Unified Evaluation across Behavioral Granularities

Model ReleasesDGX agent

arXiv:2603.02578v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in socially sensitive domains, yet their unpredictable behaviors, ranging from misalign

Intersectional Sycophancy: How Perceived User Demographics Shape False Validation in Large Language Models

Model ReleasesDGX agent

arXiv:2604.11609v1 Announce Type: new Abstract: Large language models exhibit sycophantic tendencies--validating incorrect user beliefs to appear agreeable. We investigate whether this behavior varies

Is there somewhere a collection of the best agent/coding harnesses for each models, especially open-source and local ones? In my opinion, th…

AgentsDGX agent

Is there somewhere a collection of the best agent/coding harnesses for each models, especially open-source and local ones? In my opinion, the biggest reason why people are struggling with open/local m

It is going to be like what happened in coding: as soon as models crossed a certain threshold (Opus 4.5, GPT-5.2, Gemini 3), suddenly Claude…

Model ReleasesDGX agent

It is going to be like what happened in coding: as soon as models crossed a certain threshold (Opus 4.5, GPT-5.2, Gemini 3), suddenly Claude Code & Codex were viable. Before that, it was all about cod

LAST: Leveraging Tools as Hints to Enhance Spatial Reasoning for Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2604.09712v1 Announce Type: cross Abstract: Spatial reasoning is a cornerstone capability for intelligent systems to perceive and interact with the physical world. However, multimodal large lang

let's go open-source and local models!

Model ReleasesDGX agent

let's go open-source and local models! Uber's CTO told @LauraBratton5 that AI coding tools—particularly Anthropic’s Claude Code—has already maxed out its 2026 AI budget 📈 “I'm back to the drawing boar

Like a Hammer, It Can Build, It Can Break: Large Language Model Uses, Perceptions, and Adoption in Cybersecurity Operations on Reddit

SafetyDGX agent

arXiv:2604.09998v1 Announce Type: cross Abstract: Large language models (LLMs) have recently emerged as promising tools for augmenting Security Operations Center (SOC) workflows, with vendors increasi

Locket: Robust Feature-Locking Technique for Language Models

ResearchDGX agent

arXiv:2510.12117v3 Announce Type: replace-cross Abstract: Chatbot service providers (e.g., OpenAI) rely on tiered subscription plans to generate revenue, offering black-box access to basic models for

MIXAR: Scaling Autoregressive Pixel-based Language Models to Multiple Languages and Scripts

ResearchDGX agent

arXiv:2604.11575v1 Announce Type: new Abstract: Pixel-based language models are gaining momentum as alternatives to traditional token-based approaches, promising to circumvent tokenization challenges.

Nationality encoding in language model hidden states: Probing culturally differentiated representations in persona-conditioned academic text

Model ReleasesDGX agent

arXiv:2604.10151v1 Announce Type: new Abstract: Large language models are increasingly used as writing tools and pedagogical resources in English for Academic Purposes, but it remains unclear whether

Physics-informed AI Accelerated Retention Analysis of Ferroelectric Vertical NAND: From Day-Scale TCAD to Second-Scale Surrogate Model

Model ReleasesDGX agent

arXiv:2603.06881v2 Announce Type: replace-cross Abstract: Ferroelectric field-effect transistors (FeFET)-based vertical NAND (Fe-VNAND) has emerged as a promising candidate to overcome z-scaling limit

ProGAL-VLA: Grounded Alignment through Prospective Reasoning in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2604.09824v1 Announce Type: cross Abstract: Vision language action (VLA) models enable generalist robotic agents but often exhibit language ignorance, relying on visual shortcuts and remaining i

Seg2Change: Adapting Open-Vocabulary Semantic Segmentation Model for Remote Sensing Change Detection

Model ReleasesDGX agent

arXiv:2604.11231v1 Announce Type: new Abstract: Change detection is a fundamental task in remote sensing, aiming to quantify the impacts of human activities and ecological dynamics on land-cover chang

Self-Calibrating Language Models via Test-Time Discriminative Distillation

ResearchDGX agent

arXiv:2604.09624v1 Announce Type: new Abstract: Large language models (LLMs) are systematically overconfident: they routinely express high certainty on questions they often answer incorrectly. Existin

SinkTrack: Attention Sink based Context Anchoring for Large Language Models

ResearchDGX agent

arXiv:2604.10027v1 Announce Type: new Abstract: Large language models (LLMs) suffer from hallucination and context forgetting. Prior studies suggest that attention drift is a primary cause of these pr

Stop burning GPU time on the wrong model. Manifest now supports Ollama Cloud 🦙🦚

Local AiDGX agent

Manifest is a tool or platform that has added support for Ollama Cloud, enabling users to route AI model inference to cloud-hosted Ollama instances rather than relying solely on local GPU resources. T

TokUR: Token-Level Uncertainty Estimation for Large Language Model Reasoning

ResearchDGX agent

arXiv:2505.11737v4 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) have demonstrated impressive capabilities, their output quality remains inconsistent across various applica

Towards Efficient Large Vision-Language Models: A Comprehensive Survey on Inference Strategies

ResearchDGX agent

arXiv:2603.27960v2 Announce Type: replace-cross Abstract: Although Large Vision Language Models (LVLMs) have demonstrated impressive multimodal reasoning capabilities, their scalability and deployment

Understanding Generalization in Role-Playing Models via Information Theory

ApplicationsDGX agent

arXiv:2512.17270v2 Announce Type: replace-cross Abstract: Role-playing models (RPMs) are widely used in real-world applications but underperform when deployed in the wild. This degradation can be attr

Users accuse Anthropic of degrading Claude Opus 4.6's and Claude Code's performance; the startup's employees publicly deny it degrades models to manage capacity (Carl Franzen/VentureBeat)

Model ReleasesDGX agent

Carl Franzen / VentureBeat: Users accuse Anthropic of degrading Claude Opus 4.6's and Claude Code's performance; the startup's employees publicly deny it degrades models to manage capacity — A growing

Using Deep Learning Models Pretrained by Self-Supervised Learning for Protein Localization

TutorialsDGX agent

arXiv:2604.10970v1 Announce Type: new Abstract: Background: Task-specific microscopy datasets are often small, making it difficult to train deep learning models that learn robust features. While self-

What Users Leave Unsaid: Under-Specified Queries Limit Vision-Language Models

Model ReleasesDGX agent

arXiv:2601.06165v2 Announce Type: replace-cross Abstract: Current vision-language benchmarks predominantly feature well-structured questions with clear, explicit prompts. However, real user queries ar

When simulations look right but causal effects go wrong: Large language models as behavioral simulators

SafetyDGX agent

arXiv:2604.02458v2 Announce Type: replace-cross Abstract: Behavioral simulation is increasingly used to anticipate responses to interventions. Large language models (LLMs) enable researchers to specif

YUV20K: A Complexity-Driven Benchmark and Trajectory-Aware Alignment Model for Video Camouflaged Object Detection

Model ReleasesDGX agent

arXiv:2604.09985v1 Announce Type: new Abstract: Video Camouflaged Object Detection (VCOD) is currently constrained by the scarcity of challenging benchmarks and the limited robustness of models agains

13 Apr 2026

A Predictive View on Streaming Hidden Markov Models

ResearchDGX agent

arXiv:2604.09208v1 Announce Type: cross Abstract: We develop a predictive-first optimisation framework for streaming hidden Markov models. Unlike classical approaches that prioritise full posterior re

AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention

SafetyDGX agent

arXiv:2511.18960v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have shown remarkable progress in embodied tasks recently, but most methods process visual observations in

CT-1: Vision-Language-Camera Models Transfer Spatial Reasoning Knowledge to Camera-Controllable Video Generation

TutorialsDGX agent

arXiv:2604.09201v1 Announce Type: new Abstract: Camera-controllable video generation aims to synthesize videos with flexible and physically plausible camera movements. However, existing methods either

Cybersecurity analysis: Claude Mythos Preview had a 73% success rate on expert-level capture-the-flag challenges, which no model could finish before April 2025 (AI Security Institute)

Model ReleasesDGX agent

AI Security Institute: Cybersecurity analysis: Claude Mythos Preview had a 73% success rate on expert-level capture-the-flag challenges, which no model could finish before April 2025 — The AI Security

Does anyone know which model and potentially Lora was used to create these?

Local AiDGX agent

This Reddit thread from r/StableDiffusion is a community-driven reverse-identification request, where a user shares AI-generated images and asks fellow community members to help determine which Stable

EgoTL: Egocentric Think-Aloud Chains for Long-Horizon Tasks

Model ReleasesDGX agent

arXiv:2604.09535v1 Announce Type: new Abstract: Large foundation models have made significant advances in embodied intelligence, enabling synthesis and reasoning over egocentric input for household ta

Growing a Multi-head Twig via Distillation and Reinforcement Learning to Accelerate Large Vision-Language Models

ResearchDGX agent

arXiv:2503.14075v3 Announce Type: replace-cross Abstract: Large vision-language models (VLMs) have demonstrated remarkable capabilities in open-world multimodal understanding, yet their high computati

I made a playable ping pong game where every frame is ai generated. This is my interactive diffusion model I made from scratch.

Local AiDGX agent

A Reddit user on r/StableDiffusion showcased a fully playable ping pong game in which every individual frame is rendered in real time by a custom-built interactive diffusion model, rather than using t

Inpaint workflows for z-image, qwen and flux fill onereward

Model ReleasesDGX agent

This Reddit post from r/StableDiffusion shares ComfyUI inpainting workflows for several modern AI image models, including Z-Image, Qwen Image/Edit, and Flux-series models . Flux Fill is a dedicated in

Kill-Chain Canaries: Stage-Level Tracking of Prompt Injection Across Attack Surfaces and Model Safety Tiers

Model ReleasesDGX agent

arXiv:2603.28013v3 Announce Type: replace-cross Abstract: Multi-agent LLM systems are entering production -- processing documents, managing workflows, acting on behalf of users -- yet their resilience

llama4 108b

Local AiDGX agent

This Reddit thread on r/ollama discusses running Meta's Llama 4 Maverick — a ~108B parameter model — locally using Ollama. Llama 4 models are natively multimodal AI models supporting text and image un

Mind the Gap Between Spatial Reasoning and Acting! Step-by-Step Evaluation of Agents With Spatial-Gym

TutorialsDGX agent

arXiv:2604.09338v1 Announce Type: new Abstract: Spatial reasoning is central to navigation and robotics, yet measuring model capabilities on these tasks remains difficult. Existing benchmarks evaluate

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: Multi-Exposure Image Fusion in Dynamic Scenes (Track 2)

Model ReleasesDGX agent

arXiv:2604.09030v1 Announce Type: new Abstract: This paper presents NTIRE 2026, the 3rd Restore Any Image Model (RAIM) challenge on multi-exposure image fusion in dynamic scenes. We introduce a benchm

Quantisation Reshapes the Metacognitive Geometry of Language Models

Model ReleasesDGX agent

arXiv:2604.08976v1 Announce Type: new Abstract: We report that model quantisation restructures domain-level metacognitive efficiency in LLMs rather than degrading it uniformly. Evaluating Llama-3-8B-I

Scaling model size is hitting diminishing returns. The real gains are in orchestration. Our Co-Founder & Co-CEO @yshoham makes the case in a…

Model ReleasesDGX agent

Scaling model size is hitting diminishing returns. The real gains are in orchestration. Our Co-Founder & Co-CEO @yshoham makes the case in a rare long-form profile by @Calcalistech today. The man tryi

SCoRe: Clean Image Generation from Diffusion Models Trained on Noisy Images

SafetyDGX agent

arXiv:2604.09436v1 Announce Type: new Abstract: Diffusion models trained on noisy datasets often reproduce high-frequency training artifacts, significantly degrading generation quality. To address thi

SessionIntentBench: A Multi-task Inter-session Intention-shift Modeling Benchmark for E-commerce Customer Behavior Understanding

Model ReleasesDGX agent

arXiv:2507.20185v2 Announce Type: replace Abstract: Session history is a common way of recording user interacting behaviors throughout a browsing activity with multiple products. For example, if an us

The Hot Mess of AI: How Does Misalignment Scale With Model Intelligence and Task Complexity?

SafetyDGX agent

arXiv:2601.23045v2 Announce Type: replace Abstract: As AI becomes more capable, we entrust it with more general and consequential tasks. The risks from failure grow more severe with increasing task sc

← Previous
1…121122123124125…1009
Next →