AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,597 results
14 Apr 2026

Eliciting Medical Reasoning with Knowledge-enhanced Data Synthesis: A Semi-Supervised Reinforcement Learning Approach

Model ReleasesDGX agent

arXiv:2604.11547v1 Announce Type: cross Abstract: While large language models hold promise for complex medical applications, their development is hindered by the scarcity of high-quality reasoning dat

GMI Cloud and @Zai_org are bringing fast inference + frontier models around the globe. First stop: Singapore. Big congrats to all the builde…

AgentsDGX agent

GMI Cloud and @Zai_org are bringing fast inference + frontier models around the globe. First stop: Singapore. Big congrats to all the builders at the GMI x @Zai_org Agent Hackathon in 🇸🇬 100+ on-site,

GrOCE:Graph-Guided Online Concept Erasure for Text-to-Image Diffusion Models

Research
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2511.12968v2 Announce Type: replace Abstract: Concept erasure aims to remove harmful, inappropriate, or copyrighted content from text-to-image diffusion models while preserving non-target semant

GTASA: Ground Truth Annotations for Spatiotemporal Analysis, Evaluation and Training of Video Models

SafetyDGX agent

arXiv:2604.10385v1 Announce Type: new Abstract: Generating complex multi-actor scenario videos remains difficult even for state-of-the-art neural generators, while evaluating them is hard due to the l

Long-Horizon Streaming Video Generation via Hybrid Attention with Decoupled Distillation

Local AiDGX agent

arXiv:2604.10103v1 Announce Type: new Abstract: Streaming video generation (SVG) distills a pretrained bidirectional video diffusion model into an autoregressive model equipped with sliding window att

Machine Learning-Based Detection of MCP Attacks

AgentsDGX agent

arXiv:2604.10534v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) is a new and emerging technology that extends the functionality of large language models, improving workflows but als

Mining Attribute Subspaces for Efficient Fine-tuning of 3D Foundation Models

ResearchDGX agent

arXiv:2604.10095v1 Announce Type: new Abstract: With the emergence of 3D foundation models, there is growing interest in fine-tuning them for downstream tasks, where LoRA is the dominant fine-tuning p

🦔OpenAI is backing an Illinois state bill that would shield AI labs from liability in cases where their models cause mass casualties or lar…

SafetyDGX agent

🦔OpenAI is backing an Illinois state bill that would shield AI labs from liability in cases where their models cause mass casualties or large-scale financial disasters, defined as death or serious inj

Physics-Informed State Space Models for Reliable Solar Irradiance Forecasting in Off-Grid Systems

AgentsDGX agent

arXiv:2604.11807v1 Announce Type: cross Abstract: The stable operation of autonomous off-grid photovoltaic systems dictates reliance on solar forecasting algorithms that respect atmospheric thermodyna

Prompt Injection as Role Confusion

SafetyDGX agent

arXiv:2603.12277v3 Announce Type: replace-cross Abstract: Language models remain vulnerable to prompt injection attacks despite extensive safety training. We trace this failure to role confusion: mode

Revisiting Compositionality in Dual-Encoder Vision-Language Models: The Role of Inference

SafetyDGX agent

arXiv:2604.11496v1 Announce Type: cross Abstract: Dual-encoder Vision-Language Models (VLMs) such as CLIP are often characterized as bag-of-words systems due to their poor performance on compositional

Running local models for coding — what's your actual context strategy for large codebases?

Local AiDGX agent

This r/ollama community thread discusses practical strategies for managing context windows when using locally-run LLMs (via Ollama) for coding assistance on large codebases, where context limitations

SHARE: Social-Humanities AI for Research and Education

Model ReleasesDGX agent

arXiv:2604.11152v1 Announce Type: new Abstract: This intermediate technical report introduces the SHARE family of base models and the MIRROR user interface. The SHARE models are the first causal langu

Some of my thoughts on why I loved Hampshire and why I think it should be a model for education in the AI era.

ApplicationsDGX agent

Some of my thoughts on why I loved Hampshire and why I think it should be a model for education in the AI era. Ken Burns, Lynn Pasquerella, and I on education, our rapidly changing world, and the valu

SVD-Prune: Training-Free Token Pruning For Efficient Vision-Language Models

SafetyDGX agent

arXiv:2604.11530v1 Announce Type: cross Abstract: Vision-Language Models (VLM) have revolutionized multimodal learning by jointly processing visual and textual information. Yet, they face significant

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems

Model ReleasesDGX agent

arXiv:2604.11309v1 Announce Type: cross Abstract: Large Language Models (LLMs) face prominent security risks from jailbreaking, a practice that manipulates models to bypass built-in security constrain

thinking model not working?

IndustryDGX agent

This r/ChatGPT thread addresses user-reported issues with ChatGPT's 'Thinking' mode not functioning as expected, a common frustration among subscribers. Users have noted problems such as the thinking

Training-Free Model Ensemble for Single-Image Super-Resolution via Strong-Branch Compensation

ResearchDGX agent

arXiv:2604.11564v1 Announce Type: new Abstract: Single-image super-resolution has progressed from deep convolutional baselines to stronger Transformer and state-space architectures, yet the correspond

Transformers Learn the Optimal DDPM Denoiser for Multi-Token GMMs

TutorialsDGX agent

arXiv:2604.10074v1 Announce Type: new Abstract: Transformer-based diffusion models have demonstrated remarkable performance at generating high-quality samples. However, our theoretical understanding o

Two-year-old Surface PCs get 300 price hikes as sub-1,000 models go away

IndustryDGX agent

Microsoft silently raised prices on practically all of its Surface PCs by between 100 and 500 , with the last two models previously under 1,000 — the Surface Pro 12-inch and Surface Laptop 13-inch, wh

Version numbers are not a very useful way to understand model ability gains at this stage. Unfortunately that means that if you aren’t follo…

ApplicationsDGX agent

Version numbers are not a very useful way to understand model ability gains at this stage. Unfortunately that means that if you aren’t following closely, you would expect that 5.4 is a small gain over

VLMaterial: Vision-Language Model-Based Camera-Radar Fusion for Physics-Grounded Material Identification

SafetyDGX agent

arXiv:2604.11671v1 Announce Type: cross Abstract: Accurate material recognition is a fundamental capability for intelligent perception systems to interact safely and effectively with the physical worl

WM-DAgger: Enabling Efficient Data Aggregation for Imitation Learning with World Models

SafetyDGX agent

arXiv:2604.11351v1 Announce Type: new Abstract: Imitation learning is a powerful paradigm for training robotic policies, yet its performance is limited by compounding errors: minor policy inaccuracies

13 Apr 2026

Apparently Chatgpt will end a conversation over your mom jokes. I've been begging the voice model to come back and it just ghosted me

IndustryDGX agent

A Reddit post from r/ChatGPT humorously describes a user's experience of ChatGPT's voice model abruptly ending a conversation in response to 'your mom' jokes, with the user then comically lamenting be

BEDTime: A Unified Benchmark for Automatically Describing Time Series

Model ReleasesDGX agent

arXiv:2509.05215v3 Announce Type: replace Abstract: Recent works propose complex multi-modal models that handle both time series and language, ultimately claiming high performance on complex tasks lik

deepagents subagents are just tools. when you call a subagent, thats conceptually a function call. this is the simplest mental model for bui…

AgentsDGX agent

deepagents subagents are just tools. when you call a subagent, thats conceptually a function call. this is the simplest mental model for building multiagent systems. https://docs.langchain.com/oss/pyt

Frequency-Enhanced Diffusion Models: Curriculum-Guided Semantic Alignment for Zero-Shot Skeleton Action Recognition

SafetyDGX agent

arXiv:2604.09063v1 Announce Type: cross Abstract: Human action recognition is pivotal in computer vision, with applications ranging from surveillance to human-robot interaction. Despite the effectiven

from my experience, even the best models (Opus 4.6, 5.4 xhigh / 5.3 codex) cannot write good code today without an amount of work that is eq…

TutorialsDGX agent

from my experience, even the best models (Opus 4.6, 5.4 xhigh / 5.3 codex) cannot write good code today without an amount of work that is equivalent to just doing the work myself am excited for a worl

How are you feeding personal context to your local models?

Local AiDGX agent

This r/ollama community thread discusses methods that users employ to inject personal context — such as notes, documents, and preferences — into locally-run AI models via Ollama. Common approaches exp

How Should Video LLMs Output Time? An Analysis of Efficient Temporal Grounding Paradigms

Model ReleasesDGX agent

arXiv:2604.08966v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) have advanced Video Temporal Grounding (VTG), existing methods often couple output paradigms with differe

How to build effective reward functions with AWS Lambda for Amazon Nova model customization

TutorialsDGX agent

This post demonstrates how Lambda enables scalable, cost-effective reward functions for Amazon Nova customization. You'll learn to choose between Reinforcement Learning via Verifiable Rewards (RLVR) f

LADR: Locality-Aware Dynamic Rescue for Efficient Text-to-Image Generation with Diffusion Large Language Models

Local AiDGX agent

arXiv:2603.13450v2 Announce Type: replace-cross Abstract: Discrete Diffusion Language Models have emerged as a compelling paradigm for unified multimodal generation, yet their deployment is hindered b

Large Language Models Generate Harmful Content Using a Distinct, Unified Mechanism

SafetyDGX agent

arXiv:2604.09544v1 Announce Type: cross Abstract: Large language models (LLMs) undergo alignment training to avoid harmful behaviors, yet the resulting safeguards remain brittle: jailbreaks routinely

Leave My Images Alone: Preventing Multi-Modal Large Language Models from Analyzing Images via Visual Prompt Injection

SafetyDGX agent

arXiv:2604.09024v1 Announce Type: cross Abstract: Multi-modal large language models (MLLMs) have emerged as powerful tools for analyzing Internet-scale image data, offering significant benefits but al

Multi-User Large Language Model Agents

AgentsDGX agent

arXiv:2604.08567v1 Announce Type: new Abstract: Large language models (LLMs) and LLM-based agents are increasingly deployed as assistants in planning and decision making, yet most existing systems are

Off-the-shelf Vision Models Benefit Image Manipulation Localization

Local AiDGX agent

arXiv:2604.09096v1 Announce Type: new Abstract: Image manipulation localization (IML) and general vision tasks are typically treated as two separate research directions due to the fundamental differen

Task-agnostic Low-rank Residual Adaptation for Efficient Federated Continual Fine-Tuning

Model ReleasesDGX agent

arXiv:2505.12318v2 Announce Type: replace Abstract: Federated Parameter-Efficient Fine-Tuning (Fed-PEFT) enables lightweight adaptation of large pre-trained models in federated learning settings by up

The nextAI Solution to the NeurIPS 2023 LLM Efficiency Challenge

Model ReleasesDGX agent

arXiv:2604.09034v1 Announce Type: new Abstract: The rapid evolution of Large Language Models (LLMs) has significantly impacted the field of natural language processing, but their growing complexity ra

Trained a Qwen2.5-0.5B-Instruct bf16 model on Reddit post summarization task with GRPO [P]

ResearchDGX agent

A community practitioner post on r/MachineLearning documenting an experiment fine-tuning Alibaba's Qwen2.5-0.5B-Instruct model in bf16 precision on a Reddit post summarization task using GRPO (Group R

12 Apr 2026

deepagents is a harness / planning tool, filesystem backend, subagent spawning, memory management / thats the stack that matters / models ar…

AgentsDGX agent

deepagents is a harness / planning tool, filesystem backend, subagent spawning, memory management / thats the stack that matters / models are the cpu, the harness is the os / anyways, check it out htt

Only OG's know @NousResearch had bots back in 2024. This is when models were not capable. They've tried to solve this problem every way poss…

AgentsDGX agent

Only OG's know @NousResearch had bots back in 2024. This is when models were not capable. They've tried to solve this problem every way possible. Even @karan4d was exploring such ideas acitvely, @max_

The differentiating factor between a prototype and an autonomous system is no longer solely the underlying model weights, but the sophistica…

AgentsDGX agent

The differentiating factor between a prototype and an autonomous system is no longer solely the underlying model weights, but the sophistication of the orchestration layer and its capacity for continu

11 Apr 2026

I built modern AI client for Mac with agentic tools, elegant UI, interactive charts and maps, sortable tables, Slack-like threads and access to local and cloud models

AgentsDGX agent

A Reddit post on r/ollama showcasing a community-built, feature-rich macOS AI client designed for both local and cloud model access, including support for Ollama. The application emphasizes a modern,

What are the current best models quality-wise?

Local AiDGX agent

This r/StableDiffusion thread discusses community recommendations for the highest-quality image generation models available. Flux 2 is widely regarded as arguably the best overall image generation mod

10 Apr 2026

Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models

AgentsDGX agent

arXiv:2604.08545v1 Announce Type: new Abstract: The advent of agentic multimodal models has empowered systems to actively interact with external environments. However, current agents suffer from a pro

GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents

Model ReleasesDGX agent

arXiv:2604.07429v1 Announce Type: new Abstract: Towards an embodied generalist for real-world interaction, Multimodal Large Language Model (MLLM) agents still suffer from challenging latency, sparse f

GEAR: GEometry-motion Alternating Refinement for Articulated Object Modeling with Gaussian Splatting

ResearchDGX agent

arXiv:2604.07728v1 Announce Type: new Abstract: High-fidelity interactive digital assets are essential for embodied intelligence and robotic interaction, yet articulated objects remain challenging to

How Does Machine Learning Manage Complexity?

ResearchDGX agent

arXiv:2604.07233v1 Announce Type: new Abstract: We provide a computational complexity lens to understand the power of machine learning models, particularly their ability to model complex systems. Mach

IatroBench: Pre-Registered Evidence of Iatrogenic Harm from AI Safety Measures

Model ReleasesDGX agent

arXiv:2604.07709v1 Announce Type: cross Abstract: Ask a frontier model how to taper six milligrams of alprazolam (psychiatrist retired, ten days of pills left, abrupt cessation causes seizures) and it

Large Language Models for Outpatient Referral: Problem Definition, Benchmarking and Challenges

ApplicationsDGX agent

arXiv:2503.08292v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly applied to outpatient referral tasks across healthcare systems. However, there is a lack of stan

Lost in Cultural Translation: Do LLMs Struggle with Math Across Cultural Contexts?

Model ReleasesDGX agent

arXiv:2503.18018v2 Announce Type: replace Abstract: We demonstrate that large language models' (LLMs) mathematical reasoning is culturally sensitive: testing 14 models from Anthropic, OpenAI, Google,

LumiCtrl : Learning Illuminant Prompts for Lighting Control in Personalized Text-to-Image Models

ResearchDGX agent

arXiv:2512.17489v2 Announce Type: replace Abstract: Text-to-image (T2I) models have demonstrated remarkable progress in creative image generation, yet they still lack precise control over scene illumi

MDP modeling for multi-stage stochastic programs

SafetyDGX agent

arXiv:2509.22981v2 Announce Type: replace Abstract: We study a class of multi-stage stochastic programs, which incorporate modeling features from Markov decision processes (MDPs). This class includes

OceanMAE: A Foundation Model for Ocean Remote Sensing

TutorialsDGX agent

arXiv:2604.08171v1 Announce Type: new Abstract: Accurate ocean mapping is essential for applications such as bathymetry estimation, seabed characterization, marine litter detection, and ecosystem moni

People in Washington get played, yet again We really should worry about cybersecurity - a lot – but Mythos is not the model these guys think…

SafetyDGX agent

People in Washington get played, yet again We really should worry about cybersecurity - a lot – but Mythos is not the model these guys think it is. (See my newsletter today for three reasons why it is

SeLaR: Selective Latent Reasoning in Large Language Models

ResearchDGX agent

arXiv:2604.08299v1 Announce Type: new Abstract: Chain-of-Thought (CoT) has become a cornerstone of reasoning in large language models, yet its effectiveness is constrained by the limited expressivenes

UniLACT: Depth-Aware RGB Latent Action Learning for Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2602.20231v2 Announce Type: replace-cross Abstract: Latent action representations learned from unlabeled videos have recently emerged as a promising paradigm for pretraining vision-language-acti

We had Lin on stage: 'the future is millions of models — one per application, one per use case.' Jet delivered a masterclass on reinforcemen…

ToolsDGX agent

We had Lin on stage: 'the future is millions of models — one per application, one per use case.' Jet delivered a masterclass on reinforcement fine-tuning. Rob joined @WorkOS for some hot takes on the

9 Apr 2026

maybe some nuance 😄 I don’t think anyone is “lying” about how great Mythos will be —> but there’s expectation misalignment between the Test…

Model ReleasesDGX agent

maybe some nuance 😄 I don’t think anyone is “lying” about how great Mythos will be —> but there’s expectation misalignment between the Test Harness set up for Mythos and a belief it was given this cra

Multimodal Embedding & Reranker Models with Sentence Transformers

ToolsDGX agent

The Sentence Transformers v5.4 update introduces first-class multimodal support, enabling the same familiar API to encode and compare texts, images, audio, and videos using both `SentenceTransforme...

← Previous
1…193194195196197…1010
Next →