AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,520 results
Model Releases

EviDAG: Auditable Causal DAG Authoring with Biomedical Literature

DGX agent

arXiv:2607.21859v2 Announce Type: replace Abstract: Constructing causal directed acyclic graphs (DAGs) is a core step in biomedical causal analysis, yet it remains a largely manual process. Analysts m

model-releasesarxiv-cs-ai
29 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Fast, accurate, and differentiable: a neural-network surrogate for NRSur7dq4 precessing binary black hole waveforms

DGX agent

arXiv:2607.24960v1 Announce Type: cross Abstract: We present a neural network surrogate model that emulates the NRSur7dq4 gravitational waveform model for precessing binary black hole mergers. The sur

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

Few-Shot Open-Vocabulary Remote Sensing Segmentation via Textual Inversion

DGX agent

arXiv:2607.25563v1 Announce Type: new Abstract: Open-vocabulary segmentation labels arbitrary categories from a text query without per-class training, yet on remote sensing imagery it underperforms on

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

First Kimi K3 results on home lab ~ 4t/s

DGX agent

I've got better results than expected for 768gb DDR5 and 2x5090. Using fork https://github.com/pwilkin/llama.cpp/tree/kimi-k3-text and https://huggingface.co/GrEarl/Kimi-K3-GGUF Q2_K quant. Prefill sp

model-releasesr-localllama
29 Jul 2026
Model Releases

Food Image Segmentation with LLM-Derived Ingredient Labels and Multimodal Fusion

DGX agent

arXiv:2607.25820v1 Announce Type: new Abstract: Food image segmentation plays a vital role in health-related applications such as nutrition tracking and personalized health monitoring. However, existi

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Forensic Reproducibility Audit of a Radiology Vision-Language Model Benchmark: From Intended Protocol to Released Artifact

DGX agent

arXiv:2607.25589v1 Announce Type: cross Abstract: Medical-imaging AI benchmarks combine datasets, DICOM rendering, prompts, provider APIs, automated labels, statistical code, manuscripts, and reposito

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

From Cellular Responses to Pharmacological Domains: Multimodal Zero-Shot Drug Representation Learning

DGX agent

arXiv:2607.25322v1 Announce Type: new Abstract: Multimodal drug discovery enables drug representation learning beyond chemical structure by incorporating cellular responses such as gene expression and

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

From CUDA to MLX: How K-Search Brings Decades of Kernel Expertise to Apple Silicon

DGX agent

Figure 1: CUDA-to-MLX optimization translation map. CUDA optimization knowledge can be translated into architecture-native MLX strategies rather than copied instruction-for-instruction. We face a new

model-releasesberkeley-ai-research
29 Jul 2026
Model Releases

From Deterministic to Generative Deep Learning for Urban Air Quality Reconstruction from Sparse Observations

DGX agent

arXiv:2607.25687v1 Announce Type: cross Abstract: Full-field reconstruction of air pollution is essential for evaluating pollution exposure and supporting public health decision-making. However, the c

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

GAUGE: Grading Agent-Built Financial Models Without a Golden Answer

DGX agent

arXiv:2607.24889v1 Announce Type: cross Abstract: Financial models combine public disclosures with analyst assumptions to produce forecasts and valuations. While some components can be checked mechani

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Generalization from Low- to Moderate-Resolution Spectra with Neural Networks for Stellar Parameter Estimation: A Case Study with DESI

DGX agent

arXiv:2602.15021v2 Announce Type: replace-cross Abstract: Cross-survey generalization is a critical challenge in stellar spectral analysis, particularly in cases such as transferring from low- to mode

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

GPT-5.6 Sol has been used to solve open problems in mathematics. So why was it struggling with ARC-AGI-3, a benchmark of 2D puzzle games? We…

DGX agent

GPT-5.6 Sol has been used to solve open problems in mathematics. So why was it struggling with ARC-AGI-3, a benchmark of 2D puzzle games? We investigated. The harness was not letting it remember what

model-releasesopenai--x
29 Jul 2026
Model Releases

GraphRareBench: An Auditable Graph-Evidence Benchmark for Phenotype-Driven Rare-Disease Diagnosis

DGX agent

arXiv:2607.24878v1 Announce Type: cross Abstract: Phenotype-driven diagnostic benchmarks usually report the rank of the reference disease, but they rarely reveal which plausible alternatives are ranke

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

GrocLM: Grocery Category Recommendation in E-Commerce with Large Language Models

DGX agent

arXiv:2607.24764v1 Announce Type: new Abstract: The rapid growth of online grocery shopping requires recommendation systems that capture cyclical purchasing behavior and diverse user intents. Traditio

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

HANDBOOK.md: A Benchmark for Long-Context Agentic Instruction Following

DGX agent

arXiv:2607.25398v1 Announce Type: new Abstract: Language-model agents are increasingly deployed under standing instructions: a system prompt, a policy file, or a skills document is placed in context,

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Harm is not Universal: Community-Specific Toxicity Detection is Urgently Needed

DGX agent

arXiv:2607.24898v1 Announce Type: cross Abstract: State-of-the-art toxicity detectors for text-to-image generation adopt a one-size-fits-all approach: a single universal model applying fixed safety gu

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

HOBA: Hierarchical On-Policy Bidding Agents for Adaptive Online Advertising

DGX agent

arXiv:2607.24779v1 Announce Type: new Abstract: Online advertising bidding systems typically deploy multiple offline-trained expert models (e.g., PID controllers, model predictive control, offline RL

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

How Small Can You Go? A Controlled Study of LoRA Rank, Target Modules, and Quantization Trade-offs for Text-to-SQL on a 60M-Parameter Model

DGX agent

arXiv:2607.25583v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) and low-bit quantization are now standard tools for adapting language models under tight compute budgets, yet the

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

https://x.com/EMostaque/status/2082600218529235174

DGX agent

After deployment, we applied GPT-5.6 Sol to advance the frontier of efficiency by making itself more efficient to run. The results: - 20% lower serving costs from production GPU kernel improvements. -

model-releasesemad-mostaque--x
29 Jul 2026
Model Releases

I pre-trained a 700m on 18B tokens optimized for Python and Wikitext | TheOneWhoWill/Shibai-700M-Base · Hugging Face

DGX agent

I know this is the 1000000th new sub billion parameter model out there and probably isn't as good as Qwen 3 0.6B or Qwen 3.5 0.8B but it still packs a decent punch. My intention to to continuously pre

model-releasesr-localllama
29 Jul 2026
Model Releases

I tried running a 1.56TB MoE model on a 6GB RTX 4050 Laptop, Here’s the result

DGX agent

The Test Bench Setup I tested running a massive 1.56TB Mixture-of-Experts (MoE) checkpoint (96 shards, 93 layers, 896 experts/layer, ~4.46 bits/param MXFP4) on a budget gaming laptop. Laptop: HP Victu

model-releasesr-localllama
29 Jul 2026
Model Releases

IMPRINT: Image-Conditioned Query Enrichment for Long-Tail Object Goal Navigation

DGX agent

arXiv:2607.25106v1 Announce Type: new Abstract: Embodied AI increasingly relies on queryable semantic maps built from pre-trained vision-language models to enable zero-shot Object Goal Navigation (Obj

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Influence of Prompt Engineering on Small Language Models for Guarded Query Routing

DGX agent

arXiv:2607.24801v1 Announce Type: cross Abstract: We study the problem of guarded query routing, where we assume that a user query first meets a router that either determines the ideal endpoint for in

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

Inspect India Evals: An Open Benchmarking Framework for Evaluating Large Language Models in the Indian Linguistic and Cultural Context

DGX agent

arXiv:2607.25375v1 Announce Type: new Abstract: India is a vast nation of over 1.4 billion people, varied by hundreds of diverse and locally specific traditions and cultures and 22 officially recogniz

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

Install the open-source Codex Security CLI,: npm install @OpenAI/codex-security Or start with: npx @OpenAI/codex-security@latest --help NPM:…

DGX agent

OpenAI has released the open‑source Codex Security CLI, which can be installed with `npm install @OpenAI/codex-security` or run directly via `npx @OpenAI/codex-security@latest --help`. The tool scans

model-releasesopenai--x
29 Jul 2026
Model Releases

Instruction-based Image Editing: A Survey on Data, Models, Evaluation, and Applications

DGX agent

arXiv:2607.25642v1 Announce Type: cross Abstract: Instruction-based Image Editing (IIE) aims to transform a given image into a new one based on textual instructions. Advances in Large Language Models

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

Instruction-Tuned Language Models Cannot Sample from Distributions They Can Describe

DGX agent

arXiv:2607.25292v1 Announce Type: new Abstract: Silicon sampling uses language models as proxies for human survey respondents, treating each model call as an independent draw from the persona's respon

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Instruction-Tuned Models Locally Reuse Human Syntax More Than Humans Do

DGX agent

arXiv:2607.26015v1 Announce Type: new Abstract: Syntactic convergence (the tendency of speakers to adapt in language towards the grammatical profiles of their interlocutors) is a well-documented featu

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

Interactive Reward Agent: GUI Task Evaluation via Environment-State Verification

DGX agent

arXiv:2607.25904v1 Announce Type: new Abstract: Graphical user interface task evaluation aims to determine whether a GUI agent has successfully completed a user instruction. Automated GUI task evaluat

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

JobMatchAI-An Intelligent Job Matching Platform Using Knowledge Graphs, Semantic Search and Explainable AI

DGX agent

arXiv:2603.14558v3 Announce Type: replace Abstract: Recruiters and job seekers rely on search systems to navigate labor markets, making candidate matching engines critical for hiring outcomes. Most sy

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Joint Text-Audio Alignment for EEG-to-Text Decoding in Chinese Speech Production and Perception

DGX agent

arXiv:2607.25626v1 Announce Type: new Abstract: Decoding speech information directly from scalp electroencephalography (EEG) into text provides a potential non-invasive neural communication pathway fo

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels

DGX agent

arXiv:2607.24762v1 Announce Type: new Abstract: Machine learning models are increasingly embedded in everyday software, and most of their runtime is spent in a small set of compute kernels such as mat

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

LaP-Forensics: Latent-Pixel Consistency Guided Multimodal Reasoning for Deepfake Detection

DGX agent

arXiv:2607.25962v1 Announce Type: new Abstract: Recent generative models can produce images with few obvious visual artifacts, weakening detectors and explanations that rely only on surface appearance

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Laplace-PSN-IRT: Uncertainty Quantification for Neural Item Response Theory Models of LLM Benchmarks

DGX agent

arXiv:2607.25257v1 Announce Type: cross Abstract: Item Response Theory (IRT) has recently been proposed as a framework for evaluating large language model (LLM) benchmarks by separating a model's late

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Learned, Relied Upon, or Necessary? Separating Checkpoint Dependence from Task-Level Value in Sheaf GNNs

DGX agent

arXiv:2607.25387v1 Announce Type: new Abstract: Learned restriction maps in sheaf graph neural networks are often treated as proof that the model has discovered useful edge geometry. That conclusion d

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

Learning from 53.6K Real-World Developer Edits of AI-Generated Code

DGX agent

arXiv:2607.25130v1 Announce Type: cross Abstract: Imperfections in AI-generated code require that software developers modify the generated code manually, or by re-prompting an AI programming assistant

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Less Data, Better Alignment: Data-Centric Multi-Evaluator Agreement for Preference Optimization

DGX agent

arXiv:2607.25136v1 Announce Type: new Abstract: Research on preference optimization often varies the training objective while holding the data fixed. We instead ask whether a small, high-confidence se

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Localized Adaptation Reveals Distinct Learning Signatures in Transformers

DGX agent

arXiv:2607.25663v1 Announce Type: new Abstract: Transformer adaptation is typically distributed across model depth, even when the intended change is narrow. We investigate how adaptation site shapes w

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Localized Anomaly Detection via Differentiable D-vine Copulas

DGX agent

arXiv:2607.25020v1 Announce Type: new Abstract: Vine copulas provide a flexible framework for modeling complex multivariate distributions through a hierarchical decomposition into bivariate pair-copul

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

M^2PO: Multi-Perspective Multi-Pair Preference Optimization for Machine Translation

DGX agent

arXiv:2510.13434v2 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with human preferences is pivotal for Machine Translation (MT), yet current approaches are often hindered by m

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model

DGX agent

arXiv:2607.24904v1 Announce Type: cross Abstract: Standard vision-language models (VLMs) suffer from Moravec's paradox: they excel at complex offline visual reasoning but struggle with simple streamin

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

Mechanisms of Width Scaling in Normalized Residual Networks: The Effective Alignment Dimension

DGX agent

arXiv:2607.24887v1 Announce Type: cross Abstract: Existing theories of neural-network width characterize asymptotic limits, but provide limited guidance on whether an expansion direction identified fr

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Med-R^3: Enhancing Medical Retrieval-Augmented Reasoning of LLMs via Progressive Reinforcement Learning

DGX agent

arXiv:2507.23541v5 Announce Type: replace Abstract: In medical scenarios, effectively retrieving external knowledge and leveraging it for rigorous logical reasoning is of significant importance. Despi

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing

DGX agent

arXiv:2607.25300v1 Announce Type: new Abstract: Video editing is fundamentally message-driven: even from the same source footage, the selected shots change depending on the narrative the editor wishes

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

MedJudgeRAG: Option-Wise Evidence Judgment with Dynamic Knowledge Graphs for Medical MCQA

DGX agent

arXiv:2607.24838v1 Announce Type: cross Abstract: In medical multiple-choice question answering (MCQA), Retrieval-Augmented Generation (RAG) can supplement the domain knowledge of language models (LMs

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Memory for Large Language Models

DGX agent

arXiv:2607.25380v1 Announce Type: new Abstract: Memory has evolved into a foundational architectural dimension in large language models (LLMs), shifting from an implicit byproduct of computation to a

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

MemSFT: Mitigating Alignment Tax with an External Parametric Memory

DGX agent

arXiv:2607.25614v1 Announce Type: cross Abstract: Adapting Large Language Models (LLMs) to specialized domains often incurs an alignment tax, as fine-tuning on domain-specific tasks can cause catastro

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

Messier: A High-Resolution Corpus for Cross-Benchmark Agent Evaluation

DGX agent

arXiv:2607.25891v1 Announce Type: new Abstract: Evaluating AI agents in interactive environments is hindered by fragmented tasks, scaffolds, verifiers, and scoring rules. Existing efforts focus on nar

model-releasesarxiv-cs-ai
29 Jul 2026
← Previous
1…7172737475…470
Next →