AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,522 results
23 Apr 2026

Beyond Text-Dominance: Understanding Modality Preference of Omni-modal Large Language Models

Model ReleasesDGX agent

arXiv:2604.16902v2 Announce Type: replace Abstract: Native Omni-modal Large Language Models (OLLMs) have shifted from pipeline architectures to unified representation spaces. However, this native inte

By end of next year. End of year after for 16gb vram. Do folk really need more for most things? Not really. But just call the cloud models w…

Local AiDGX agent

By end of next year. End of year after for 16gb vram. Do folk really need more for most things? Not really. But just call the cloud models when you do. What's a reasonable timeline for a GPT 5.4/Opus

Cold-Start Forecasting of New Product Life-Cycles via Conditional Diffusion Models

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.20370v1 Announce Type: new Abstract: Forecasting the life-cycle trajectory of a newly launched product is important for launch planning, resource allocation, and early risk assessment. This

Depression Risk Assessment in Social Media via Large Language Models

ResearchDGX agent

arXiv:2604.19887v1 Announce Type: cross Abstract: Depression is one of the most prevalent and debilitating mental health conditions worldwide, frequently underdiagnosed and undertreated. The prolifera

Development and Preliminary Evaluation of a Domain-Specific Large Language Model for Tuberculosis Care in South Africa

Model ReleasesDGX agent

arXiv:2604.19776v1 Announce Type: new Abstract: Tuberculosis (TB) is one of the world's deadliest infectious diseases, and in South Africa, it contributes a significant burden to the country's health

Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models

Model ReleasesDGX agent

arXiv:2511.06209v4 Announce Type: replace Abstract: LLMs can solve complex tasks by generating long, multi-step reasoning chains. Test-time scaling (TTS) can further improve LLM performance by samplin

Expert Upcycling: Shifting the Compute-Efficient Frontier of Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2604.19835v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become the dominant architecture for scaling large language models: frontier models routinely decouple total parameters f

GPT-5.5 in Codex is a delight to work with: - Super sharp with responses - It understands intent better than any model - Great 'personality'…

Model ReleasesDGX agent

GPT-5.5 in Codex is a delight to work with: - Super sharp with responses - It understands intent better than any model - Great 'personality' - Gets lots of stuff done without pausing unnecessarily It

How close are models to building a perfect Slack clone with less than $50,000 of tokens? I feel like not that far...

Model ReleasesDGX agent

How close are models to building a perfect Slack clone with less than 50,000 of tokens? I feel like not that far... Claude Code spend had gotten to 10.95M runrate peak at SemiAnalysis But then Opus 4.

In ChatGPT, full-stack inference improvements enable a more capable model at faster speed. This efficiency is a game-changer for GPT-5.5 Pro…

Model ReleasesDGX agent

In ChatGPT, full-stack inference improvements enable a more capable model at faster speed. This efficiency is a game-changer for GPT-5.5 Pro, now a much more practical option for demanding tasks, and

Infection-Reasoner: A Compact Vision-Language Model for Wound Infection Classification with Evidence-Grounded Clinical Reasoning

Model ReleasesDGX agent

arXiv:2604.19937v1 Announce Type: cross Abstract: Assessing chronic wound infection from photographs is challenging because visual appearance varies across wound etiologies, anatomical locations, and

Kimi K2.6, the new state-of-the-art open-weight model from Moonshot, is now available for Pro and Max subscribers.

ToolsDGX agent

Moonshot has released Kimi K2.6, an open-weight language model that represents a new advancement in the field, now made available to Pro and Max tier subscribers on Perplexity. The model's open-weight

KOCO-BENCH: Can Large Language Models Leverage Domain Knowledge in Software Development?

Model ReleasesDGX agent

arXiv:2601.13240v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at general programming but struggle with domain-specific software development, necessitating domain special

Large Language Models Outperform Humans in Fraud Detection and Resistance to Motivated Investor Pressure

Model ReleasesDGX agent

arXiv:2604.20652v1 Announce Type: new Abstract: Large language models trained on human feedback may suppress fraud warnings when investors arrive already persuaded of a fraudulent opportunity. We test

Machine learning moment closure models for the radiative transfer equation IV: enforcing symmetrizable hyperbolicity in two dimensions

ResearchDGX agent

arXiv:2604.20143v1 Announce Type: cross Abstract: This is our fourth work in the series on machine learning (ML) moment closure models for the radiative transfer equation (RTE). In the first three pap

MasconCube: Fast and Accurate Gravity Modeling with an Explicit Representation

ResearchDGX agent

arXiv:2509.08607v3 Announce Type: replace-cross Abstract: The geodesy of irregularly shaped small bodies presents fundamental challenges for gravitational field modeling, particularly as deep space ex

Mechanistic Interpretability Tool for AI Weather Models

ResearchDGX agent

arXiv:2604.20467v1 Announce Type: cross Abstract: Artificial Intelligence (AI) weather models are improving rapidly, and their forecasts are already competitive with long-established traditional Numer

Mitigating Hallucinations in Large Vision-Language Models without Performance Degradation

Model ReleasesDGX agent

arXiv:2604.20366v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) exhibit powerful generative capabilities but frequently produce hallucinations that compromise output reliability.

So right now basically anyone with a 16GB VRAM card can go on @huggingface and download a model which BEATS Claude Sonnet 4.5, all running l…

Model ReleasesDGX agent

So right now basically anyone with a 16GB VRAM card can go on @huggingface and download a model which BEATS Claude Sonnet 4.5, all running locally 🤯 LOOK AT THE WATER PARTICLES?? What are they doing i

Spatio-temporal modelling of electric vehicle charging demand

Model ReleasesDGX agent

arXiv:2604.19841v1 Announce Type: cross Abstract: Accurate forecasting of electric vehicle (EV) charging demand is critical for grid management and infrastructure planning. Yet the field continues to

Stabilising Generative Models of Attitude Change

ResearchDGX agent

arXiv:2604.19791v1 Announce Type: new Abstract: Attitude change - the process by which individuals revise their evaluative stances - has been explained by a set of influential but competing verbal the

Synthetic Flight Data Generation Using Generative Models

ApplicationsDGX agent

arXiv:2604.20293v1 Announce Type: new Abstract: The increasing adoption of synthetic data in aviation research offers a promising solution to data scarcity and confidentiality challenges. This study i

Text Steganography with Dynamic Codebook and Multimodal Large Language Model

ResearchDGX agent

arXiv:2604.20269v1 Announce Type: cross Abstract: With the popularity of the large language models (LLMs), text steganography has achieved remarkable performance. However, existing methods still have

The GaoYao Benchmark: A Comprehensive Framework for Evaluating Multilingual and Multicultural Abilities of Large Language Models

Model ReleasesDGX agent

arXiv:2604.20225v1 Announce Type: new Abstract: Evaluating the multilingual and multicultural capabilities of Large Language Models (LLMs) is essential for their global utility. However, current bench

The Ratchet Effect in Silico through Interaction-Driven Cumulative Intelligence in Large Language Models

Model ReleasesDGX agent

arXiv:2507.21166v2 Announce Type: replace-cross Abstract: Human intelligence scales through cumulative cultural evolution (CCE), a ratchet process in which innovations are retained against entropic dr

Toward Safe Autonomous Robotic Endovascular Interventions using World Models

Model ReleasesDGX agent

arXiv:2604.20151v1 Announce Type: cross Abstract: Autonomous mechanical thrombectomy (MT) presents substantial challenges due to highly variable vascular geometries and the requirements for accurate,

Towards reconstructing experimental sparse-view X-ray CT data with diffusion models

ApplicationsDGX agent

arXiv:2602.12755v2 Announce Type: replace Abstract: Diffusion-based image generators are promising priors for ill-posed inverse problems like sparse-view X-ray Computed Tomography (CT). As most studie

WISCA: A Lightweight Model Transition Method to Improve LLM Training via Weight Scaling

ResearchDGX agent

arXiv:2508.16676v2 Announce Type: replace-cross Abstract: Transformer architecture gradually dominates the LLM field. Recent advances in training optimization for Transformer-based large language mode

22 Apr 2026

Agent-GWO: Collaborative Agents for Dynamic Prompt Optimization in Large Language Models

Model ReleasesDGX agent

arXiv:2604.18612v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in complex reasoning tasks, while recent prompting strategies such as Chain-of-Thou

Alibaba launches Qwen3.6-27B, an open-weight dense model with 27B parameters, saying it surpasses Qwen3.5-397B-A17B on major coding benchmarks (Qwen)

Model ReleasesDGX agent

Qwen: Alibaba launches Qwen3.6-27B, an open-weight dense model with 27B parameters, saying it surpasses Qwen3.5-397B-A17B on major coding benchmarks — · 4226 words · QwenTeam丨Translations:.体中文 — HUGGI

All of the AI models have preferred names. If you asked Claude 4.5 for a software developer, you are going to get Marcus Chen. Wizards are m…

Model ReleasesDGX agent

All of the AI models have preferred names. If you asked Claude 4.5 for a software developer, you are going to get Marcus Chen. Wizards are mostly named Aldric. Space pilots are Kira from Claude, Mara

Bayesian Event-Based Model for Disease Subtype and Stage Inference

ApplicationsDGX agent

arXiv:2512.03467v2 Announce Type: replace Abstract: Chronic diseases often progress differently across patients. Rather than randomly varying, there are typically a small number of subtypes for how a

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models

SafetyDGX agent

arXiv:2509.26238v4 Announce Type: replace Abstract: Monitoring large language models' (LLMs) activations is an effective way to detect harmful requests before they lead to unsafe outputs. However, tra

Bridging the High-Frequency Data Gap: A Millisecond-Resolution Network Dataset for Advancing Time Series Foundation Models

ApplicationsDGX agent

arXiv:2603.16497v2 Announce Type: replace-cross Abstract: Time series foundation models (TSFMs) require diverse, real-world datasets to adapt across varying domains and temporal frequencies. However,

CASS: Nvidia to AMD Transpilation with Data, Models, and Benchmark

Model ReleasesDGX agent

arXiv:2505.16968v4 Announce Type: replace-cross Abstract: Cross-architecture GPU code transpilation is essential for unlocking low-level hardware portability, yet no scalable solution exists. We intro

Comparing energy consumption and accuracy in text classification inference

ResearchDGX agent

arXiv:2508.14170v2 Announce Type: replace Abstract: The increasing deployment of large language models (LLMs) in natural language processing (NLP) tasks raises concerns about energy efficiency and sus

Concept Inconsistency in Dermoscopic Concept Bottleneck Models: A Rough-Set Analysis of the Derm7pt Dataset

Model ReleasesDGX agent

arXiv:2604.19323v1 Announce Type: cross Abstract: Concept Bottleneck Models (CBMs) route predictions exclusively through a clinically grounded concept layer, binding interpretability to concept-label

ConvVitMamba: Efficient Multiscale Convolution, Transformer, and Mamba-Based Sequence modelling for Hyperspectral Image Classification

Model ReleasesDGX agent

arXiv:2604.18856v1 Announce Type: new Abstract: Hyperspectral image (HSI) classification remains challenging due to high spectral dimensionality, redundancy, and limited labeled data. Although convolu

Diff-SBSR: Learning Multimodal Feature-Enhanced Diffusion Models for Zero-Shot Sketch-Based 3D Shape Retrieval

SafetyDGX agent

arXiv:2604.19135v1 Announce Type: new Abstract: This paper presents the first exploration of text-to-image diffusion models for zero-shot sketch-based 3D shape retrieval (ZS-SBSR). Existing sketch-bas

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling

SafetyDGX agent

arXiv:2604.19544v1 Announce Type: new Abstract: Multimodal reward models (MRMs) play a crucial role in aligning Multimodal Large Language Models (MLLMs) with human preferences. Training a good MRM req

DUALVISION: RGB-Infrared Multimodal Large Language Models for Robust Visual Reasoning

Model ReleasesDGX agent

arXiv:2604.18829v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved impressive performance on visual perception and reasoning tasks with RGB imagery, yet they remain

FairTree: Subgroup Fairness Auditing of Machine Learning Models with Bias-Variance Decomposition

SafetyDGX agent

arXiv:2604.19357v1 Announce Type: new Abstract: The evaluation of machine learning models typically relies mainly on performance metrics based on loss functions, which risk to overlook changes in perf

Hierarchically Robust Zero-shot Vision-language Models

SafetyDGX agent

arXiv:2604.18867v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) can perform zero-shot classification but are susceptible to adversarial attacks. While robust fine-tuning improves their

How to Teach Large Multimodal Models New Skills

SafetyDGX agent

arXiv:2510.08564v2 Announce Type: replace Abstract: How can we teach large multimodal models (LMMs) new skills without erasing prior abilities? We study sequential fine-tuning on five target skills wh

InHabit: Leveraging Image Foundation Models for Scalable 3D Human Placement

ApplicationsDGX agent

arXiv:2604.19673v1 Announce Type: new Abstract: Training embodied agents to understand 3D scenes as humans do requires large-scale data of people meaningfully interacting with diverse environments, ye

Introducing Gemini Enterprise Agent Platform, powering the next wave of agents

Model ReleasesDGX agent

In the early days of generative AI, building safe and reliable business tools took massive engineering effort and a high tolerance for trial and error. We helped solve that with Vertex AI, our trusted

LBLLM: Lightweight Binarization of Large Language Models via Three-Stage Distillation

HardwareDGX agent

arXiv:2604.19167v1 Announce Type: cross Abstract: Deploying large language models (LLMs) in resource-constrained environments is hindered by heavy computational and memory requirements. We present LBL

Local Linearity of LLMs Enables Activation Steering via Model-Based Linear Optimal Control

Local AiDGX agent

arXiv:2604.19018v1 Announce Type: cross Abstract: Inference-time LLM alignment methods, particularly activation steering, offer an alternative to fine-tuning by directly modifying activations during g

OLLM: Options-based Large Language Models

Model ReleasesDGX agent

arXiv:2604.19087v1 Announce Type: new Abstract: We introduce Options LLM (OLLM), a simple, general method that replaces the single next-token prediction of standard LLMs with a extit{set of learned op

OmniParser V2: Structured-Points-of-Thought for Unified Visual Text Parsing and Its Generality to Multimodal Large Language Models

ResearchDGX agent

arXiv:2502.16161v2 Announce Type: replace-cross Abstract: Visually-situated text parsing (VsTP) has recently seen notable advancements, driven by the growing demand for automated document understandin

Once Correct, Still Wrong: Counterfactual Hallucination in Multilingual Vision-Language Models

Model ReleasesDGX agent

arXiv:2602.05437v2 Announce Type: replace Abstract: Vision-language models (VLMs) can achieve high accuracy while still accepting culturally plausible but visually incorrect interpretations. Existing

ParamBoost: Gradient Boosted Piecewise Cubic Polynomials

ApplicationsDGX agent

arXiv:2604.18864v1 Announce Type: new Abstract: Generalized Additive Models (GAMs) can be used to create non-linear glass-box (i.e. explicitly interpretable) models, where the predictive function is f

Presumably GPT-imagegen-2 (aka ChatGPT Images 2.0 aka gpt-image-2) works as a tool which the models generate prompts for? I wish we could se…

Model ReleasesDGX agent

Presumably GPT-imagegen-2 (aka ChatGPT Images 2.0 aka gpt-image-2) works as a tool which the models generate prompts for? I wish we could see those prompts, like back in the DALL-E 3 days https://simo

ProjLens: Unveiling the Role of Projectors in Multimodal Model Safety

SafetyDGX agent

arXiv:2604.19083v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success in cross-modal understanding and generation, yet their deployment is threate

Q-Mask: Query-driven Causal Masks for Text Anchoring in OCR-Oriented Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.00161v2 Announce Type: replace Abstract: Optical Character Recognition (OCR) is increasingly regarded as a foundational capability for modern vision-language models (VLMs), enabling them no

R^2-dLLM: Accelerating Diffusion Large Language Models via Spatio-Temporal Redundancy Reduction

Local AiDGX agent

arXiv:2604.18995v1 Announce Type: cross Abstract: Diffusion Large Language Models (dLLMs) have emerged as a promising alternative to autoregressive generation by enabling parallel token prediction. Ho

Rethinking Scale: Deployment Trade-offs of Small Language Models under Agent Paradigms

AgentsDGX agent

arXiv:2604.19299v1 Announce Type: cross Abstract: Despite the impressive capabilities of large language models, their substantial computational costs, latency, and privacy risks hinder their widesprea

Safety-Critical Contextual Control via Online Riemannian Optimization with World Models

SafetyDGX agent

arXiv:2604.19639v1 Announce Type: cross Abstract: Modern world models are becoming too complex to admit explicit dynamical descriptions. We study safety-critical contextual control, where a Planner mu

Sources: OpenAI has been briefing US federal agencies, state governments, and Five Eyes allies on the capabilities of its GPT-5.4-Cyber model over the past week (Sam Sabin/Axios)

Model ReleasesDGX agent

Sam Sabin / Axios: Sources: OpenAI has been briefing US federal agencies, state governments, and Five Eyes allies on the capabilities of its GPT-5.4-Cyber model over the past week — OpenAI has been br

TEMPO: Scaling Test-time Training for Large Reasoning Models

SafetyDGX agent

arXiv:2604.19295v1 Announce Type: new Abstract: Test-time training (TTT) adapts model parameters on unlabeled test instances during inference time, which continuously extends capabilities beyond the r

← Previous
1…118119120121122…1009
Next →