AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,563 results
Model Releases

AlphaInventory: Evolving White-Box Inventory Policies via Large Language Models with Deployment Guarantees

DGX agent

arXiv:2605.00369v1 Announce Type: new Abstract: We study how large language models can be used to evolve inventory policies in online, non-stationary environments. Our work is motivated by recent adva

model-releasesarxiv-cs-lg
4 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

April 2026 newsletter

DGX agent

I just sent out the April edition of my sponsors-only monthly newsletter. If you are a sponsor (or if you start a sponsorship now) you can access it here. In this month's newsletter: Opus 4.7 and GPT-

model-releasessimon-willison
4 May 2026
Model Releases

At any point in time, you can safely resume to using Anthropic's models: ollama launch claude-desktop --restore

DGX agent

This post discusses Ollama's functionality for resuming work with Anthropic's Claude models, indicating that users can safely restore previous sessions or states using a command-line interface (`ollam

model-releasesollama--x
4 May 2026
Model Releases

BanglaSocialBench: A Benchmark for Evaluating Sociopragmatic and Cultural Alignment of LLMs in Bangladeshi Social Interaction

DGX agent

arXiv:2603.15949v3 Announce Type: replace Abstract: Large Language Models have demonstrated strong multilingual fluency, yet fluency alone does not guarantee socially appropriate language use. In high

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Beyond Benchmarks: MathArena as an Evaluation Platform for Mathematics with LLMs

DGX agent

arXiv:2605.00674v1 Announce Type: new Abstract: Large language models (LLMs) are becoming increasingly capable mathematical collaborators, but static benchmarks are no longer sufficient for evaluating

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Beyond Heuristics: Learnable Density Control for 3D Gaussian Splatting

DGX agent

arXiv:2605.00408v1 Announce Type: new Abstract: While 3D Gaussian Splatting (3DGS) has demonstrated impressive real-time rendering performance, its efficacy remains constrained by a reliance on heuris

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Beyond Visual Fidelity: Benchmarking Super-Resolution Models for Large-Scale Remote Sensing Imagery via Downstream Task Integration

DGX agent

arXiv:2605.00310v1 Announce Type: new Abstract: Super-resolution (SR) techniques have made major advances in reconstructing high-resolution images from low-resolution inputs. The increased resolution

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Borrowed Geometry: Computational Reuse of Frozen Text-Pretrained Transformer Weights Across Modalities

DGX agent

arXiv:2605.00333v1 Announce Type: cross Abstract: Frozen Gemma 4 31B weights pretrained exclusively on text tokens, unmodified, transfer across modality boundaries through a thin trainable interface.

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Bring Your Own Prompts: Use-Case-Specific Bias and Fairness Evaluation for LLMs

DGX agent

arXiv:2407.10853v5 Announce Type: replace Abstract: Bias and fairness risks in Large Language Models (LLMs) vary substantially across deployment contexts, yet existing approaches lack systematic guida

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Can Coding Agents Reproduce Findings in Computational Materials Science?

DGX agent

arXiv:2605.00803v1 Announce Type: cross Abstract: Large language models are increasingly deployed as autonomous coding agents and have achieved remarkably strong performance on software engineering be

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Caracal: Causal Architecture via Spectral Mixing

DGX agent

arXiv:2605.00292v1 Announce Type: new Abstract: The scalability of Large Language Models to long sequences is hindered by the quadratic cost of attention and the limitations of positional encodings. T

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

claudely: launch Claude Code against Local LLM provider like LM Studio / Ollama / llama.cpp without trashing your real claude config

DGX agent

claudely is a tool that enables users to run Claude Code against local LLM providers such as LM Studio, Ollama, or llama.cpp while preserving their existing Claude configuration. The tool allows devel

model-releasesr-ollama
4 May 2026
Model Releases

CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection

DGX agent

arXiv:2605.00630v1 Announce Type: new Abstract: The proliferation of advanced AI video synthesis techniques poses an unprecedented challenge to digital video authenticity. Existing AI-generated video

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

CompleteRXN: Toward Completing Open Chemical Reaction Databases

DGX agent

arXiv:2605.00222v1 Announce Type: new Abstract: Chemical reaction datasets such as USPTO suffer from substantial incompleteness, frequently missing byproducts, co-reactants, and stoichiometric coeffic

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Concolic Testing on Individual Fairness of Neural Network Models

DGX agent

arXiv:2509.06864v2 Announce Type: replace Abstract: This paper introduces PyFair, a formal framework for evaluating and verifying individual fairness of Deep Neural Networks (DNNs). By adapting the co

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

ControBench: An Interaction-Aware Benchmark for Controversial Discourse Analysis on Social Networks

DGX agent

arXiv:2605.00513v1 Announce Type: new Abstract: Understanding how people argue across ideological divides online is important for studying political polarization, misinformation, and content moderatio

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

CURE-OOD: Benchmarking Out-of-Distribution Detection for Survival Prediction

DGX agent

arXiv:2605.00350v1 Announce Type: new Abstract: ``How long can I live and remain free of cancer?'' is often the first question a patient asks after receiving a cancer diagnosis and treatment. Accurate

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

deepagents-cli is quietly becoming the best place to start coding with open weight models. we've been investing heavily in making it a harne…

DGX agent

deepagents-cli is quietly becoming the best place to start coding with open weight models. we've been investing heavily in making it a harness that's truly model-agnostic, without compromising perform

model-releasesharrison-chase--x
4 May 2026
Model Releases

Deepfakes: we need to re-think the concept of 'real' images

DGX agent

arXiv:2509.21864v2 Announce Type: replace Abstract: The wide availability and low usability barrier of modern image generation models has triggered the reasonable fear of criminal misconduct and negat

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Deepseek V4 works more thoroughly than other open source models: It writes its own tests and performs extensive validation. This leads to be…

DGX agent

Deepseek V4 works more thoroughly than other open source models: It writes its own tests and performs extensive validation. This leads to better performance, but also cases of the model being overconf

model-releasesjeremy-howard--x
4 May 2026
Model Releases

Differentiable Autoencoding Neural Operator for Interpretable and Integrable Latent Space Modeling

DGX agent

arXiv:2510.00233v2 Announce Type: replace Abstract: Scientific machine learning has enabled the extraction of physical insights and data-driven modeling of high-dimensional spatiotemporal data, yet ac

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Do Open-Loop Metrics Predict Closed-Loop Driving? A Cross-Benchmark Correlation Study of NAVSIM and Bench2Drive

DGX agent

arXiv:2605.00066v1 Announce Type: new Abstract: Open-loop evaluation offers fast, reproducible assessment of autonomous driving planners, but its ability to predict real closed-loop driving performanc

model-releasesarxiv-cs-ro
4 May 2026
Model Releases

Documentation https://docs.ollama.com/integrations/claude-desktop

DGX agent

This documentation page describes how to integrate Ollama with Claude Desktop, enabling users to run local language models through the Anthropic Claude interface. The integration allows Claude Desktop

model-releasesollama--x
4 May 2026
Model Releases

Driving with A Thousand Faces: A Benchmark for Closed-Loop Personalized End-to-End Autonomous Driving

DGX agent

arXiv:2602.18757v2 Announce Type: replace Abstract: Human driving behavior is inherently diverse, yet most end-to-end autonomous driving (E2E-AD) systems learn a single average driving style, neglecti

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Dynamics-Encoded Deep Learning for Robust System Identification and Parameter Estimation

DGX agent

arXiv:2410.04299v2 Announce Type: replace Abstract: Incorporating a priori physics knowledge into machine learning leads to more robust and interpretable algorithms. In this work, we combine deep lear

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Efficient Spatio-Temporal Vegetation Pixel Classification with Vision Transformers

DGX agent

arXiv:2605.00296v1 Announce Type: new Abstract: Plant phenology-the study of recurrent life cycle events-is essential for understanding ecosystem dynamics and their responses to climate change impacts

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Evaluating the Architectural Reasoning Capabilities of LLM Provers via the Obfuscated Natural Number Game

DGX agent

arXiv:2605.00677v1 Announce Type: new Abstract: While Large Language Models have achieved notable success on formal mathematics benchmarks such as MiniF2F, it remains unclear whether these results ste

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Event-based Civil Infrastructure Visual Defect Detection: ev-CIVIL Dataset and Benchmark

DGX agent

arXiv:2504.05679v2 Announce Type: replace Abstract: Small unmanned aerial vehicle (UAV)-based visual inspections are a more efficient alternative to manual methods for examining civil structural defec

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Ex-iRobot CEO Colin Angle launches Familiar Machines & Magic and unveils Familiar, a dog-like, 'emotionally intelligent' robot that reacts to owner's feelings (Christopher Mims/Wall Street Journal)

DGX agent

Christopher Mims / Wall Street Journal: Ex-iRobot CEO Colin Angle launches Familiar Machines & Magic and unveils Familiar, a dog-like, “emotionally intelligent” robot that reacts to owner's feelings —

model-releasestechmeme
4 May 2026
Model Releases

ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation

DGX agent

arXiv:2507.14201v3 Announce Type: replace-cross Abstract: We present ExCyTIn-Bench, the first benchmark to Evaluate an LLM agent X on the task of Cyber Threat Investigation through security questions

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Exploring the System 1 Thinking Capability of Large Reasoning Models

DGX agent

arXiv:2504.10368v4 Announce Type: replace Abstract: This paper explores the system 1 thinking capability of Large Reasoning Models (LRMs), the intuitive ability to respond efficiently with minimal tok

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Faithful Extreme Image Rescaling with Learnable Reversible Transformation and Semantic Priors

DGX agent

arXiv:2605.00605v1 Announce Type: new Abstract: Most recent extreme rescaling methods struggle to preserve semantically consistent structures and produce realistic details, due to the severely ill-pos

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

FedACT: Concurrent Federated Intelligence across Heterogeneous Data Sources

DGX agent

arXiv:2605.00011v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative intelligence across decentralized data source devices in a privacy-preserving way. While substantial resea

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

FinSafetyBench: Evaluating LLM Safety in Real-World Financial Scenarios

DGX agent

arXiv:2605.00706v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied in financial scenarios. However, they may produce harmful outputs, including facilitating illegal

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Firestore at Next '26: Unlock agentic development, search and MongoDB compatibility

DGX agent

In the era of AI agents, the distance between a big idea and a working application has never been shorter. As we lean more heavily on agents to help us build applications, a critical question remains:

model-releasesgoogle-cloud-ai
4 May 2026
Model Releases

FollowTable: A Benchmark for Instruction-Following Table Retrieval

DGX agent

arXiv:2605.00400v1 Announce Type: cross Abstract: Table Retrieval (TR) has traditionally been formulated as an ad-hoc retrieval problem, where relevance is primarily determined by topical semantic sim

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Foresight Arena: An On-Chain Benchmark for Evaluating AI Forecasting Agents

DGX agent

arXiv:2605.00420v1 Announce Type: cross Abstract: Evaluating the true forecasting ability of AI agents requires environments resistant to overfitting, free from centralized trust, and grounded in ince

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

From Backward Spreading to Forward Replay: Revisiting Target Construction in LLM Parameter Editing

DGX agent

arXiv:2605.00358v1 Announce Type: new Abstract: LLM parameter editing methods commonly rely on computing an ideal target hidden-state at a target layer (referred as anchor point) and distributing the

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

From Images2Mesh: A 3D Surface Reconstruction Pipeline for Non-Cooperative Space Objects

DGX agent

arXiv:2605.00147v1 Announce Type: new Abstract: On-orbit inspection imagery is crucial as it enables characterization of non-cooperative resident space objects, providing the geometry and structural c

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

From Prediction to Practice: A Task-Aware Evaluation Framework for Blood Glucose Forecasting

DGX agent

arXiv:2605.00645v1 Announce Type: new Abstract: Clinical time-series forecasting is increasingly studied for decision support, yet standard aggregate metrics can obscure whether a model is actually us

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

GitHub: ComfyUI SenseNova U1 Released – Anyone Got It Working Yet for ComfyUI?

DGX agent

SenseNova U1 is a new series of native multimodal models that unifies multimodal understanding, reasoning, and generation within a monolithic architecture, marking a fundamental paradigm shift from mo

model-releasesr-stablediffusion
4 May 2026
Model Releases

Granite 4.1 3B SVG Pelican Gallery

DGX agent

Granite 4.1 3B SVG Pelican Gallery IBM released their Granite 4.1 family of LLMs a few days ago. They're Apache 2.0 licensed and come in 3B, 8B and 30B sizes. Granite 4.1 LLMs: How They’re Built by Gr

model-releasessimon-willison
4 May 2026
Model Releases

Grok 4.3 just built this entire game with just a single prompt It has the fastest output token speed and outranks Claude Sonnet 4.6 Max on A…

DGX agent

Grok 4.3 just built this entire game with just a single prompt It has the fastest output token speed and outranks Claude Sonnet 4.6 Max on Artificial Analysis I built this using the xAI API in Kilo Co

model-releaseselon-musk--x
4 May 2026
Model Releases

High-Probability Convergence in Decentralized Stochastic Optimization with Gradient Tracking

DGX agent

arXiv:2605.00281v1 Announce Type: new Abstract: We study high-probability (HP) convergence guarantees in decentralized stochastic optimization, where multiple agents collaborate to jointly train a mod

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

How Frontier LLMs Adapt to Neurodivergence Context: A Measurement Framework for Surface vs. Structural Change in System-Prompted Responses

DGX agent

arXiv:2605.00113v1 Announce Type: new Abstract: We examine if frontier chat-based large language models (LLMs) adjust their outputs based on neurodivergence (ND) context in system prompts and describe

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

How I built a free, local AI powerhouse in 10 days (Ollama + Gemma 4 + Claude Cowork 3P + Browserless)

DGX agent

This post documents a 10-day project to build a local AI system using open-source tools and models, specifically combining Ollama (a local LLM framework), Gemma 4 (a language model), Claude Cowork 3P,

model-releasesr-ollama
4 May 2026
Model Releases

How OpenAI delivers low-latency voice AI at scale

DGX agent

OpenAI describes its technical approach to delivering real-time voice AI services with minimal latency across large user bases, likely covering infrastructure optimization, model serving strategies, a

model-releasesopenai
4 May 2026
Model Releases

How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks

DGX agent

arXiv:2507.01955v3 Announce Type: replace Abstract: Multimodal foundation models (MFMs), such as GPT-4o, have recently made remarkable progress. However, their detailed visual understanding beyond que

model-releasesarxiv-cs-cv
4 May 2026
← Previous
1…364365366367368…471
Next →