AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,904 results
15 Apr 2026

On Efficient Variants of Segment Anything Model: A Survey

ApplicationsDGX agent

arXiv:2410.04960v5 Announce Type: replace Abstract: The Segment Anything Model (SAM) is a foundational model for image segmentation tasks, known for its strong generalization across diverse applicatio

Parcae: Doing more with fewer parameters using stable looped models

ToolsDGX agent

Parcae is a stable looped language model that matches the quality of a Transformer twice its size — a 770M model reaching 1.3B-level performance. We introduce the first scaling laws for looping and sh

PromptEcho: Annotation-Free Reward from Vision-Language Models for Text-to-Image Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.12652v1 Announce Type: cross Abstract: Reinforcement learning (RL) can improve the prompt following capability of text-to-image (T2I) models, yet obtaining high-quality reward signals remai

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Reading Between the Pixels: Linking Text-Image Embedding Alignment to Typographic Attack Success on Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.12371v1 Announce Type: new Abstract: We study typographic prompt injection attacks on vision-language models (VLMs), where adversarial text is rendered as images to bypass safety mechanisms

Red Teaming Large Reasoning Models

Model ReleasesDGX agent

arXiv:2512.00412v4 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs) have emerged as a powerful advancement in multi-step reasoning tasks, offering enhanced transparency and logical

RePAIR: Interactive Machine Unlearning through Prompt-Aware Model Repair

Local AiDGX agent

arXiv:2604.12820v1 Announce Type: new Abstract: Large language models (LLMs) inherently absorb harmful knowledge, misinformation, and personal data during pretraining on large-scale web corpora, with

T2I-BiasBench: A Multi-Metric Framework for Auditing Demographic and Cultural Bias in Text-to-Image Models

Model ReleasesDGX agent

arXiv:2604.12481v1 Announce Type: new Abstract: Text-to-image (T2I) generative models achieve impressive visual fidelity but inherit and amplify demographic imbalances and cultural biases embedded in

When Reasoning Models Hurt Behavioral Simulation: A Solver-Sampler Mismatch in Multi-Agent LLM Negotiation

Model ReleasesDGX agent

arXiv:2604.11840v1 Announce Type: new Abstract: Large language models are increasingly used as agents in social, economic, and policy simulations. A common assumption is that stronger reasoning should

14 Apr 2026

[13 Apr 2026] Top Local Models List - April 2026 https://www.latent.space/p/ainews-top-local-models-list-april we did the research so you do…

ToolsDGX agent

As of April 2026, Latent Space published a curated ranking of the top locally-runnable AI models, providing a research-based overview of the best open-weight or self-hostable models available at that

A collaborative agent with two lightweight synergistic models for autonomous crystal materials research

AgentsDGX agent

arXiv:2604.11540v1 Announce Type: new Abstract: Current large language models require hundreds of billions of parameters yet struggle with domain-specific reasoning and tool coordination in materials

A Minimal Mathematical Model for Conducting Patterns

Model ReleasesDGX agent

arXiv:2604.10356v1 Announce Type: cross Abstract: We present a minimal mathematical model for conducting patterns that separates geometric trajectory from temporal parametrization. The model is based

Advancing Polish Language Modeling through Tokenizer Optimization in the Bielik v3 7B and 11B Series

Model ReleasesDGX agent

arXiv:2604.10799v1 Announce Type: cross Abstract: The development of the Bielik v3 PL series, encompassing both the 7B and 11B parameter variants, represents a significant milestone in the field of la

Advancing Reasoning in Diffusion Language Models with Denoising Process Rewards

TutorialsDGX agent

arXiv:2510.01544v2 Announce Type: replace Abstract: Diffusion-based large language models offer a non-autoregressive alternative for text generation, but enabling them to perform complex reasoning rem

Back to the Barn with LLAMAs: Evolving Pretrained LLM Backbones in Finetuning Vision Language Models

Model ReleasesDGX agent

arXiv:2604.10985v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have rapidly advanced by leveraging powerful pre-trained Large Language Models (LLMs) as core reasoning backbones. As new

Do Thought Streams Matter? Evaluating Reasoning in Gemini Vision-Language Models for Video Scene Understanding

Model ReleasesDGX agent

arXiv:2604.11177v1 Announce Type: new Abstract: We benchmark how internal reasoning traces, which we call thought streams, affect video scene understanding in vision-language models. Using four config

Does Your VFM Speak Plant? The Botanical Grammar of Vision Foundation Models for Object Detection

ApplicationsDGX agent

arXiv:2604.09920v1 Announce Type: new Abstract: Vision foundation models (VFMs) offer the promise of zero-shot object detection without task-specific training data, yet their performance in complex ag

Downloading an AI model just to hit an OOM error is the worst. 📉

Local AiDGX agent

This Reddit post from r/ollama discusses the frustrating experience of spending time downloading a large AI model via Ollama only to encounter an Out-of-Memory (OOM) error when attempting to run it, m

Fairboard: a quantitative framework for equity assessment of healthcare models

Local AiDGX agent

arXiv:2604.09656v1 Announce Type: cross Abstract: Despite there now being more than 1,000 FDA-authorised AI medical devices, formal equity assessments -- whether model performance is uniform across pa

FlowHijack: A Dynamics-Aware Backdoor Attack on Flow-Matching Vision-Language-Action Models

ResearchDGX agent

arXiv:2604.09651v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are emerging as a cornerstone for robotics, with flow-matching policies like pi_0 showing great promise in generatin

Generation-Augmented Generation: A Plug-and-Play Framework for Private Knowledge Injection in Large Language Models

Model ReleasesDGX agent

arXiv:2601.08209v3 Announce Type: replace Abstract: In domains such as materials science, biomedicine, and finance, high-stakes deployment of large language models (LLMs) requires injecting private, d

Grounded World Model for Semantically Generalizable Planning

Model ReleasesDGX agent

arXiv:2604.11751v1 Announce Type: cross Abstract: In Model Predictive Control (MPC), world models predict the future outcomes of various action proposals, which are then scored to guide the selection

HiEdit: Lifelong Model Editing with Hierarchical Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.11214v1 Announce Type: new Abstract: Lifelong model editing (LME) aims to sequentially rectify outdated or inaccurate knowledge in deployed LLMs while minimizing side effects on unrelated i

How Alignment Routes: Localizing, Scaling, and Controlling Policy Circuits in Language Models

Model ReleasesDGX agent

arXiv:2604.04385v3 Announce Type: replace-cross Abstract: This paper localizes the policy routing mechanism in alignment-trained language models. An intermediate-layer attention gate reads detected co

HumorGen: Cognitive Synergy for Humor Generation in Large Language Models via Persona-Based Distillation

Model ReleasesDGX agent

arXiv:2604.09629v1 Announce Type: new Abstract: Humor generation poses a significant challenge for Large Language Models (LLMs), because their standard training objective - predicting the most likely

Immunizing 3D Gaussian Generative Models Against Unauthorized Fine-Tuning via Attribute-Space Traps

Model ReleasesDGX agent

arXiv:2604.09688v1 Announce Type: new Abstract: Recent large-scale generative models enable high-quality 3D synthesis. However, the public accessibility of pre-trained weights introduces a critical vu

Minimizing classical resources in variational measurement-based quantum computation for generative modeling

Model ReleasesDGX agent

arXiv:2604.11578v1 Announce Type: cross Abstract: Measurement-based quantum computation (MBQC) is a framework for quantum information processing in which a computational task is carried out through on

MoveFM-R: Advancing Mobility Foundation Models via Language-driven Semantic Reasoning

ResearchDGX agent

arXiv:2509.22403v2 Announce Type: replace Abstract: Mobility Foundation Models (MFMs) have advanced the modeling of human movement patterns, yet they face a ceiling due to limitations in data scale an

Neural Generalized Mixed-Effects Models

Model ReleasesDGX agent

arXiv:2604.10976v1 Announce Type: cross Abstract: Generalized linear mixed-effects models (GLMMs) are widely used to analyze grouped and hierarchical data. In a GLMM, each response is assumed to follo

Risk Awareness Injection: Calibrating Vision-Language Models for Safety without Compromising Utility

SafetyDGX agent

arXiv:2602.03402v3 Announce Type: replace Abstract: Vision language models (VLMs) extend the reasoning capabilities of large language models (LLMs) to cross-modal settings, yet remain highly vulnerabl

RobustMedSAM: Degradation-Resilient Medical Image Segmentation via Robust Foundation Model Adaptation

Model ReleasesDGX agent

arXiv:2604.09814v1 Announce Type: new Abstract: Medical image segmentation models built on Segment Anything Model (SAM) achieve strong performance on clean benchmarks, yet their reliability often degr

SCITUNE: Aligning Large Language Models with Human-Curated Scientific Multimodal Instructions

Model ReleasesDGX agent

arXiv:2307.01139v2 Announce Type: replace-cross Abstract: Instruction finetuning is a popular paradigm to align large language models (LLM) with human intent. Despite its popularity, this idea is less

VeriInteresting: An Empirical Study of Model Prompt Interactions in Verilog Code Generation

Model ReleasesDGX agent

arXiv:2603.08715v2 Announce Type: replace-cross Abstract: Rapid advances in language models (LMs) have created new opportunities for automated code generation while complicating trade-offs between mod

WaveMoE: A Wavelet-Enhanced Mixture-of-Experts Foundation Model for Time Series Forecasting

Model ReleasesDGX agent

arXiv:2604.10544v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) have recently achieved remarkable success in universal forecasting by leveraging large-scale pretraining on dive

Why Do Large Language Models Generate Harmful Content?

ResearchDGX agent

arXiv:2604.11663v1 Announce Type: new Abstract: Large Language Models (LLMs) have been shown to generate harmful content. However, the underlying causes of such behavior remain under explored. We prop

13 Apr 2026

A Representation-Level Assessment of Bias Mitigation in Foundation Models

SafetyDGX agent

arXiv:2604.08561v1 Announce Type: new Abstract: We investigate how successful bias mitigation reshapes the embedding space of encoder-only and decoder-only foundation models, offering an internal audi

Attention-Based Sampler for Diffusion Language Models

ResearchDGX agent

arXiv:2604.08564v1 Announce Type: new Abstract: Auto-regressive models (ARMs) have established a dominant paradigm in language modeling. However, their strictly sequential decoding paradigm imposes fu

Benchmarking CNN- and Transformer-Based Models for Surgical Instrument Segmentation in Robotic-Assisted Surgery

Model ReleasesDGX agent

arXiv:2604.09151v1 Announce Type: new Abstract: Accurate segmentation of surgical instruments in robotic-assisted surgery is critical for enabling context-aware computer-assisted interventions, such a

Do Vision Language Models Need to Process Image Tokens?

ResearchDGX agent

arXiv:2604.09425v1 Announce Type: new Abstract: Vision Language Models (VLMs) have achieved remarkable success by integrating visual encoders with large language models (LLMs). While VLMs process dens

Improving Automatic Summarization of Radiology Reports through Mid-Training of Large Language Models

Model ReleasesDGX agent

arXiv:2603.19275v2 Announce Type: replace-cross Abstract: Automatic summarization of radiology reports is an essential application to reduce the burden on physicians. Previous studies have widely used

11 Apr 2026

if you've read software history everything in your bone tells you open models have to win we are just in a weird anthropic fanboy moment

AgentsDGX agent

if you've read software history everything in your bone tells you open models have to win we are just in a weird anthropic fanboy moment we're seeing that open source models are getting good at file o

Why is Wan 2.2 N.S.F.W Remix Lightning Model so much better at things like hair flip, hair combing and feminine energy than regular Wan?

Local AiDGX agent

The Wan 2.2 NSFW Remix Lightning Model is a fine-tuned variant of the base Wan 2.2 video generation model, distinguished by its blending of open-source motion LoRA data and refined pose training for e

10 Apr 2026

AHCQ-SAM: Toward Accurate and Hardware-Compatible Post-Training Segment Anything Model Quantization

Model ReleasesDGX agent

arXiv:2503.03088v4 Announce Type: replace-cross Abstract: The Segment Anything Model (SAM) has revolutionized image and video segmentation with its powerful zero-shot capabilities. However, its massiv

Anthropic tries to keep its new AI model away from cyberattackers as enterprises look to tame AI chaos

Model ReleasesDGX agent

Sure, at some point quantum computing may break data encryption — but well before that, artificial intelligence models already seem likely to wreak havoc. That became starkly apparent this week when A

Emotion Concepts and their Function in a Large Language Model

Model ReleasesDGX agent

arXiv:2604.07729v1 Announce Type: cross Abstract: Large language models (LLMs) sometimes appear to exhibit emotional reactions. We investigate why this is the case in Claude Sonnet 4.5 and explore imp

Entropy After </Think> for reasoning model early exiting

Model ReleasesDGX agent

arXiv:2509.26522v3 Announce Type: replace Abstract: Reasoning LLMs show improved performance with longer chains of thought. However, recent work has highlighted their tendency to overthink, continuing

FlowGuard: Towards Lightweight In-Generation Safety Detection for Diffusion Models via Linear Latent Decoding

Model ReleasesDGX agent

arXiv:2604.07879v1 Announce Type: new Abstract: Diffusion-based image generation models have advanced rapidly but pose a safety risk due to their potential to generate Not-Safe-For-Work (NSFW) content

From Classical Machine Learning to Tabular Foundation Models: An Empirical Investigation of Robustness and Scalability Under Class Imbalance in Emergency and Critical Care

ResearchDGX agent

arXiv:2512.21602v2 Announce Type: replace-cross Abstract: Millions of patients pass through emergency departments and intensive care units each year, where clinicians must make high-stakes decisions u

I strongly suspect that Claude Mythos is a looped language model, as described in the paper 'Scaling Latent Reasoning via Looped Language Mo…

Model ReleasesDGX agent

I strongly suspect that Claude Mythos is a looped language model, as described in the paper 'Scaling Latent Reasoning via Looped Language Models' from ByteDance The authors of that paper called out gr

Illocutionary Explanation Planning for Source-Faithful Explanations in Retrieval-Augmented Language Models

Model ReleasesDGX agent

arXiv:2604.06211v1 Announce Type: cross Abstract: Natural language explanations produced by large language models (LLMs) are often persuasive, but not necessarily scrutable: users cannot easily verify

MM-MoralBench: A MultiModal Moral Evaluation Benchmark for Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2412.20718v2 Announce Type: replace Abstract: The rapid integration of Large Vision-Language Models (LVLMs) into critical domains necessitates comprehensive moral evaluation to ensure their alig

TalkLoRA: Communication-Aware Mixture of Low-Rank Adaptation for Large Language Models

Model ReleasesDGX agent

arXiv:2604.06291v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) enables parameter-efficient fine-tuning of Large Language Models (LLMs), and recent Mixture-of-Experts (MoE) extensions fur

TOOLCAD: Exploring Tool-Using Large Language Models in Text-to-CAD Generation with Reinforcement Learning

AgentsDGX agent

arXiv:2604.07960v1 Announce Type: cross Abstract: Computer-Aided Design (CAD) is an expert-level task that relies on long-horizon reasoning and coherent modeling actions. Large Language Models (LLMs)

9 Apr 2026

Use the Same Model Across Ollama, LM Studio, Jan, and your Favorite Local AI Apps

Local AiDGX agent

Local AI tools such as Ollama, LM Studio, and Jan all rely on the same underlying inference engine (llama.cpp) and support compatible model formats (primarily GGUF), meaning a single downloaded mod...

8 Apr 2026

[AINews] Meta Superintelligence Labs announces Muse Spark, first frontier model on their completely new stack

ToolsDGX agent

Muse Spark is the first model from Meta's Superintelligence Labs division and the debut entry in the new Muse model family, representing a 'ground-up overhaul' of the company's AI efforts. It is ...

Only a few hundred Tesla Model S & X cars left in inventory. Order now if you want one.

IndustryDGX agent

In April 2026, Elon Musk announced via X that only a few hundred Tesla Model S and Model X units remained in inventory, urging prospective buyers to order immediately. After more than a decade on ...

7 Apr 2026

I spent the night testing open-source coding models against Claude Opus in production. Same codebase. Same tasks. Real API calls, real file …

Model ReleasesDGX agent

I spent the night testing open-source coding models against Claude Opus in production. Same codebase. Same tasks. Real API calls, real file edits, real bugs. Tested: Arcee Trinity-Large-Thinking, http

13 Aug 2026

BEST-KAG: Enhancing Question Answering of Building Engineering Standards with Multimodal Knowledge Graph Modeling and Large Language Model

SafetyDGX agent

arXiv:2608.11244v1 Announce Type: new Abstract: Construction standards are critical for building safety and sustainability. Existing standard application workflows rely on keyword-based document retri

Do You See What You Draw? A Semantic Closed-Loop Framework for Holistic Evaluation of Unified Multimodal Models

Model ReleasesDGX agent

arXiv:2608.11907v1 Announce Type: cross Abstract: As Large Vision-Language Models increasingly aim to integrate visual generation and understanding within a single parameter space, evaluating such str

Locating and Controlling Implicit Personalization in Large Language Models

Model ReleasesDGX agent

arXiv:2608.11735v1 Announce Type: cross Abstract: Large language models (LLMs) often shift their outputs in response to implicit demographic cues even when users never state a demographic identity. Pr

Our most intelligent workhorse model yet for coding and agents has arrived ⚡ Meet Gemini 3.7 Flash. — Crush that seemingly endless to-do lis…

Model ReleasesDGX agent

Our most intelligent workhorse model yet for coding and agents has arrived ⚡ Meet Gemini 3.7 Flash. — Crush that seemingly endless to-do list. Gemini Spark in the @geminiapp now uses 3.7 Flash. The ne

← Previous
1…4344454647…999
Next →