AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,904 results
Applications

On Efficient Variants of Segment Anything Model: A Survey

DGX agent

arXiv:2410.04960v5 Announce Type: replace Abstract: The Segment Anything Model (SAM) is a foundational model for image segmentation tasks, known for its strong generalization across diverse applicatio

applicationsarxiv-cs-cv
15 Apr 2026
Tools

Parcae: Doing more with fewer parameters using stable looped models

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

Parcae is a stable looped language model that matches the quality of a Transformer twice its size — a 770M model reaching 1.3B-level performance. We introduce the first scaling laws for looping and sh

toolstogether-ai-blog
15 Apr 2026
Model Releases

PromptEcho: Annotation-Free Reward from Vision-Language Models for Text-to-Image Reinforcement Learning

DGX agent

arXiv:2604.12652v1 Announce Type: cross Abstract: Reinforcement learning (RL) can improve the prompt following capability of text-to-image (T2I) models, yet obtaining high-quality reward signals remai

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Reading Between the Pixels: Linking Text-Image Embedding Alignment to Typographic Attack Success on Vision-Language Models

DGX agent

arXiv:2604.12371v1 Announce Type: new Abstract: We study typographic prompt injection attacks on vision-language models (VLMs), where adversarial text is rendered as images to bypass safety mechanisms

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Red Teaming Large Reasoning Models

DGX agent

arXiv:2512.00412v4 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs) have emerged as a powerful advancement in multi-step reasoning tasks, offering enhanced transparency and logical

model-releasesarxiv-cs-ai
15 Apr 2026
Local Ai

RePAIR: Interactive Machine Unlearning through Prompt-Aware Model Repair

DGX agent

arXiv:2604.12820v1 Announce Type: new Abstract: Large language models (LLMs) inherently absorb harmful knowledge, misinformation, and personal data during pretraining on large-scale web corpora, with

local-aiarxiv-cs-ai
15 Apr 2026
Model Releases

T2I-BiasBench: A Multi-Metric Framework for Auditing Demographic and Cultural Bias in Text-to-Image Models

DGX agent

arXiv:2604.12481v1 Announce Type: new Abstract: Text-to-image (T2I) generative models achieve impressive visual fidelity but inherit and amplify demographic imbalances and cultural biases embedded in

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

When Reasoning Models Hurt Behavioral Simulation: A Solver-Sampler Mismatch in Multi-Agent LLM Negotiation

DGX agent

arXiv:2604.11840v1 Announce Type: new Abstract: Large language models are increasingly used as agents in social, economic, and policy simulations. A common assumption is that stronger reasoning should

model-releasesarxiv-cs-lg
15 Apr 2026
Tools

[13 Apr 2026] Top Local Models List - April 2026 https://www.latent.space/p/ainews-top-local-models-list-april we did the research so you do…

DGX agent

As of April 2026, Latent Space published a curated ranking of the top locally-runnable AI models, providing a research-based overview of the best open-weight or self-hostable models available at that

toolsswyx--x
14 Apr 2026
Agents

A collaborative agent with two lightweight synergistic models for autonomous crystal materials research

DGX agent

arXiv:2604.11540v1 Announce Type: new Abstract: Current large language models require hundreds of billions of parameters yet struggle with domain-specific reasoning and tool coordination in materials

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

A Minimal Mathematical Model for Conducting Patterns

DGX agent

arXiv:2604.10356v1 Announce Type: cross Abstract: We present a minimal mathematical model for conducting patterns that separates geometric trajectory from temporal parametrization. The model is based

model-releasesarxiv-cs-ro
14 Apr 2026
Model Releases

Advancing Polish Language Modeling through Tokenizer Optimization in the Bielik v3 7B and 11B Series

DGX agent

arXiv:2604.10799v1 Announce Type: cross Abstract: The development of the Bielik v3 PL series, encompassing both the 7B and 11B parameter variants, represents a significant milestone in the field of la

model-releasesarxiv-cs-ai
14 Apr 2026
Tutorials

Advancing Reasoning in Diffusion Language Models with Denoising Process Rewards

DGX agent

arXiv:2510.01544v2 Announce Type: replace Abstract: Diffusion-based large language models offer a non-autoregressive alternative for text generation, but enabling them to perform complex reasoning rem

tutorialsarxiv-cs-ai
14 Apr 2026
Model Releases

Back to the Barn with LLAMAs: Evolving Pretrained LLM Backbones in Finetuning Vision Language Models

DGX agent

arXiv:2604.10985v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have rapidly advanced by leveraging powerful pre-trained Large Language Models (LLMs) as core reasoning backbones. As new

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Do Thought Streams Matter? Evaluating Reasoning in Gemini Vision-Language Models for Video Scene Understanding

DGX agent

arXiv:2604.11177v1 Announce Type: new Abstract: We benchmark how internal reasoning traces, which we call thought streams, affect video scene understanding in vision-language models. Using four config

model-releasesarxiv-cs-cv
14 Apr 2026
Applications

Does Your VFM Speak Plant? The Botanical Grammar of Vision Foundation Models for Object Detection

DGX agent

arXiv:2604.09920v1 Announce Type: new Abstract: Vision foundation models (VFMs) offer the promise of zero-shot object detection without task-specific training data, yet their performance in complex ag

applicationsarxiv-cs-cv
14 Apr 2026
Local Ai

Downloading an AI model just to hit an OOM error is the worst. 📉

DGX agent

This Reddit post from r/ollama discusses the frustrating experience of spending time downloading a large AI model via Ollama only to encounter an Out-of-Memory (OOM) error when attempting to run it, m

local-air-ollama
14 Apr 2026
Local Ai

Fairboard: a quantitative framework for equity assessment of healthcare models

DGX agent

arXiv:2604.09656v1 Announce Type: cross Abstract: Despite there now being more than 1,000 FDA-authorised AI medical devices, formal equity assessments -- whether model performance is uniform across pa

local-aiarxiv-cs-ai
14 Apr 2026
Research

FlowHijack: A Dynamics-Aware Backdoor Attack on Flow-Matching Vision-Language-Action Models

DGX agent

arXiv:2604.09651v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are emerging as a cornerstone for robotics, with flow-matching policies like pi_0 showing great promise in generatin

researcharxiv-cs-cv
14 Apr 2026
Model Releases

Generation-Augmented Generation: A Plug-and-Play Framework for Private Knowledge Injection in Large Language Models

DGX agent

arXiv:2601.08209v3 Announce Type: replace Abstract: In domains such as materials science, biomedicine, and finance, high-stakes deployment of large language models (LLMs) requires injecting private, d

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Grounded World Model for Semantically Generalizable Planning

DGX agent

arXiv:2604.11751v1 Announce Type: cross Abstract: In Model Predictive Control (MPC), world models predict the future outcomes of various action proposals, which are then scored to guide the selection

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

HiEdit: Lifelong Model Editing with Hierarchical Reinforcement Learning

DGX agent

arXiv:2604.11214v1 Announce Type: new Abstract: Lifelong model editing (LME) aims to sequentially rectify outdated or inaccurate knowledge in deployed LLMs while minimizing side effects on unrelated i

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

How Alignment Routes: Localizing, Scaling, and Controlling Policy Circuits in Language Models

DGX agent

arXiv:2604.04385v3 Announce Type: replace-cross Abstract: This paper localizes the policy routing mechanism in alignment-trained language models. An intermediate-layer attention gate reads detected co

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

HumorGen: Cognitive Synergy for Humor Generation in Large Language Models via Persona-Based Distillation

DGX agent

arXiv:2604.09629v1 Announce Type: new Abstract: Humor generation poses a significant challenge for Large Language Models (LLMs), because their standard training objective - predicting the most likely

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Immunizing 3D Gaussian Generative Models Against Unauthorized Fine-Tuning via Attribute-Space Traps

DGX agent

arXiv:2604.09688v1 Announce Type: new Abstract: Recent large-scale generative models enable high-quality 3D synthesis. However, the public accessibility of pre-trained weights introduces a critical vu

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Minimizing classical resources in variational measurement-based quantum computation for generative modeling

DGX agent

arXiv:2604.11578v1 Announce Type: cross Abstract: Measurement-based quantum computation (MBQC) is a framework for quantum information processing in which a computational task is carried out through on

model-releasesarxiv-cs-ai
14 Apr 2026
Research

MoveFM-R: Advancing Mobility Foundation Models via Language-driven Semantic Reasoning

DGX agent

arXiv:2509.22403v2 Announce Type: replace Abstract: Mobility Foundation Models (MFMs) have advanced the modeling of human movement patterns, yet they face a ceiling due to limitations in data scale an

researcharxiv-cs-lg
14 Apr 2026
Model Releases

Neural Generalized Mixed-Effects Models

DGX agent

arXiv:2604.10976v1 Announce Type: cross Abstract: Generalized linear mixed-effects models (GLMMs) are widely used to analyze grouped and hierarchical data. In a GLMM, each response is assumed to follo

model-releasesarxiv-cs-lg
14 Apr 2026
Safety

Risk Awareness Injection: Calibrating Vision-Language Models for Safety without Compromising Utility

DGX agent

arXiv:2602.03402v3 Announce Type: replace Abstract: Vision language models (VLMs) extend the reasoning capabilities of large language models (LLMs) to cross-modal settings, yet remain highly vulnerabl

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

RobustMedSAM: Degradation-Resilient Medical Image Segmentation via Robust Foundation Model Adaptation

DGX agent

arXiv:2604.09814v1 Announce Type: new Abstract: Medical image segmentation models built on Segment Anything Model (SAM) achieve strong performance on clean benchmarks, yet their reliability often degr

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

SCITUNE: Aligning Large Language Models with Human-Curated Scientific Multimodal Instructions

DGX agent

arXiv:2307.01139v2 Announce Type: replace-cross Abstract: Instruction finetuning is a popular paradigm to align large language models (LLM) with human intent. Despite its popularity, this idea is less

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

VeriInteresting: An Empirical Study of Model Prompt Interactions in Verilog Code Generation

DGX agent

arXiv:2603.08715v2 Announce Type: replace-cross Abstract: Rapid advances in language models (LMs) have created new opportunities for automated code generation while complicating trade-offs between mod

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

WaveMoE: A Wavelet-Enhanced Mixture-of-Experts Foundation Model for Time Series Forecasting

DGX agent

arXiv:2604.10544v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) have recently achieved remarkable success in universal forecasting by leveraging large-scale pretraining on dive

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Why Do Large Language Models Generate Harmful Content?

DGX agent

arXiv:2604.11663v1 Announce Type: new Abstract: Large Language Models (LLMs) have been shown to generate harmful content. However, the underlying causes of such behavior remain under explored. We prop

researcharxiv-cs-ai
14 Apr 2026
Safety

A Representation-Level Assessment of Bias Mitigation in Foundation Models

DGX agent

arXiv:2604.08561v1 Announce Type: new Abstract: We investigate how successful bias mitigation reshapes the embedding space of encoder-only and decoder-only foundation models, offering an internal audi

safetyarxiv-cs-cl
13 Apr 2026
Research

Attention-Based Sampler for Diffusion Language Models

DGX agent

arXiv:2604.08564v1 Announce Type: new Abstract: Auto-regressive models (ARMs) have established a dominant paradigm in language modeling. However, their strictly sequential decoding paradigm imposes fu

researcharxiv-cs-cl
13 Apr 2026
Model Releases

Benchmarking CNN- and Transformer-Based Models for Surgical Instrument Segmentation in Robotic-Assisted Surgery

DGX agent

arXiv:2604.09151v1 Announce Type: new Abstract: Accurate segmentation of surgical instruments in robotic-assisted surgery is critical for enabling context-aware computer-assisted interventions, such a

model-releasesarxiv-cs-cv
13 Apr 2026
Research

Do Vision Language Models Need to Process Image Tokens?

DGX agent

arXiv:2604.09425v1 Announce Type: new Abstract: Vision Language Models (VLMs) have achieved remarkable success by integrating visual encoders with large language models (LLMs). While VLMs process dens

researcharxiv-cs-cv
13 Apr 2026
Model Releases

Improving Automatic Summarization of Radiology Reports through Mid-Training of Large Language Models

DGX agent

arXiv:2603.19275v2 Announce Type: replace-cross Abstract: Automatic summarization of radiology reports is an essential application to reduce the burden on physicians. Previous studies have widely used

model-releasesarxiv-cs-ai
13 Apr 2026
Agents

if you've read software history everything in your bone tells you open models have to win we are just in a weird anthropic fanboy moment

DGX agent

if you've read software history everything in your bone tells you open models have to win we are just in a weird anthropic fanboy moment we're seeing that open source models are getting good at file o

agentsharrison-chase--x
11 Apr 2026
Local Ai

Why is Wan 2.2 N.S.F.W Remix Lightning Model so much better at things like hair flip, hair combing and feminine energy than regular Wan?

DGX agent

The Wan 2.2 NSFW Remix Lightning Model is a fine-tuned variant of the base Wan 2.2 video generation model, distinguished by its blending of open-source motion LoRA data and refined pose training for e

local-air-stablediffusion
11 Apr 2026
Model Releases

AHCQ-SAM: Toward Accurate and Hardware-Compatible Post-Training Segment Anything Model Quantization

DGX agent

arXiv:2503.03088v4 Announce Type: replace-cross Abstract: The Segment Anything Model (SAM) has revolutionized image and video segmentation with its powerful zero-shot capabilities. However, its massiv

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Anthropic tries to keep its new AI model away from cyberattackers as enterprises look to tame AI chaos

DGX agent

Sure, at some point quantum computing may break data encryption — but well before that, artificial intelligence models already seem likely to wreak havoc. That became starkly apparent this week when A

model-releasessiliconangle
10 Apr 2026
Model Releases

Emotion Concepts and their Function in a Large Language Model

DGX agent

arXiv:2604.07729v1 Announce Type: cross Abstract: Large language models (LLMs) sometimes appear to exhibit emotional reactions. We investigate why this is the case in Claude Sonnet 4.5 and explore imp

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Entropy After </Think> for reasoning model early exiting

DGX agent

arXiv:2509.26522v3 Announce Type: replace Abstract: Reasoning LLMs show improved performance with longer chains of thought. However, recent work has highlighted their tendency to overthink, continuing

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

FlowGuard: Towards Lightweight In-Generation Safety Detection for Diffusion Models via Linear Latent Decoding

DGX agent

arXiv:2604.07879v1 Announce Type: new Abstract: Diffusion-based image generation models have advanced rapidly but pose a safety risk due to their potential to generate Not-Safe-For-Work (NSFW) content

model-releasesarxiv-cs-cv
10 Apr 2026
Research

From Classical Machine Learning to Tabular Foundation Models: An Empirical Investigation of Robustness and Scalability Under Class Imbalance in Emergency and Critical Care

DGX agent

arXiv:2512.21602v2 Announce Type: replace-cross Abstract: Millions of patients pass through emergency departments and intensive care units each year, where clinicians must make high-stakes decisions u

researcharxiv-cs-cv
10 Apr 2026
Model Releases

I strongly suspect that Claude Mythos is a looped language model, as described in the paper 'Scaling Latent Reasoning via Looped Language Mo…

DGX agent

I strongly suspect that Claude Mythos is a looped language model, as described in the paper 'Scaling Latent Reasoning via Looped Language Models' from ByteDance The authors of that paper called out gr

model-releasesjeremy-howard--x
10 Apr 2026
← Previous
1…5455565758…1248
Next →