AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,499 results
Model Releases

Knowledge before Reasoning: EC-Reason-Bench, a Training-Free Diagnostic Benchmark for LLM Enzyme Classification

DGX agent

arXiv:2607.26397v1 Announce Type: new Abstract: Enzyme function prediction is a hierarchical, knowledge-intensive form of protein function classification. Existing benchmarks expose an anomaly: genera

model-releasesarxiv-cs-cl
30 Jul 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Language Models are not Equally Robust to Non-Canonical Tokenization across Languages

DGX agent

arXiv:2607.26831v1 Announce Type: new Abstract: Despite the existence of exponentially many valid tokenizations for a given string, language models operate on a single canonical sequence deterministic

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Level, Sharpness, and Corpus: Why Zero-Shot OOD Detector Rankings Do Not Transfer

DGX agent

arXiv:2607.26582v1 Announce Type: new Abstract: Selecting a zero-shot out-of-distribution (OOD) detector for a new deployment is typically based on benchmark rankings, implicitly assuming that the hig

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

LG AI Research releases K-EXAONE 2.0 750B A37B

DGX agent

It was developed under Phase 2 of Korea's Sovereign AI Foundation Model Project. ​Size: 750B parameters (3x larger than their 236B v1 model). ​- License: Apache 2.0 ​Languages: Expanded to 10 language

model-releasesr-localllama
30 Jul 2026
Model Releases

llm 0.32rc1

DGX agent

Release: llm 0.32rc1 This RC for LLM 0.32 finishes the work that started in LLM 0.32a0 - it adds a new schema design that does a much better job of capturing the details of the prompts and responses r

model-releasessimon-willison
30 Jul 2026
Model Releases

llm 0.32rc2

DGX agent

Release: llm 0.32rc2 Hot on the heels of RC1, this fixes a dependency issue and also adds two neat new features: The default model for users who have not set their own default is now GPT-5.6 Luna. It

model-releasessimon-willison
30 Jul 2026
Model Releases

llm-chat-completions-server 0.1a0

DGX agent

Release: llm-chat-completions-server 0.1a0 A key goal of the new content-addressable logs in LLM 0.32rc1 was being able to support OpenAI Chat Completion style requests where each incoming message ext

model-releasessimon-willison
30 Jul 2026
Model Releases

LLMET: Enabling Cross-Layer Evaluation of Emerging M3D Memories for Energy-Efficient LLM Serving

DGX agent

arXiv:2607.26491v1 Announce Type: cross Abstract: The energy consumption of Large Language Model (LLM) serving is becoming a major system challenge as deployment scales, driven by hardware power and t

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Low-cost Embedded Breathing Rate Determination Using 802.15.4z IR-UWB Hardware for Remote Healthcare

DGX agent

arXiv:2504.03772v3 Announce Type: replace-cross Abstract: Respiratory diseases account for a significant portion of global mortality. Affordable and early detection is an effective way of addressing t

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Making advanced intelligence more abundant and affordable is central to our mission to ensure AGI benefits all of humanity. With the help of…

DGX agent

Making advanced intelligence more abundant and affordable is central to our mission to ensure AGI benefits all of humanity. With the help of GPT-5.6 Sol, we have made leaps in efficiency. Today, we ar

model-releasesopenai--x
30 Jul 2026
Model Releases

Mechanistic interpretability streamlined for everyday users like us😎 🧠

DGX agent

Context: I want to give the community an Open Research (well open under Apache 2.0 clause) - tool that allows everyday users like us to look deeper into the local models we use consistently. Mechanist

model-releasesr-localllama
30 Jul 2026
Model Releases

MediaWiki Code2Code Search: Neural Retrieval for the Semantic Discovery of Open-Source Software Entities

DGX agent

arXiv:2607.26766v1 Announce Type: cross Abstract: Code search in large-scale ecosystems is often hindered by the lexical gap between user queries and implementation details, alongside the trade-off be

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

MEDIC: Comprehensive Evaluation of Leading Indicators for LLM Safety and Utility in Clinical Applications

DGX agent

arXiv:2409.07314v4 Announce Type: replace Abstract: While Large Language Models (LLMs) achieve superhuman performance on standardized medical licensing exams, these static benchmarks have become satur

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Memory bandwidth, not VRAM size, sets your tokens/sec — here's the arithmetic

DGX agent

Every week someone asks which card to buy and the thread turns into people naming GPUs they happen to own. There's an actual calculation behind it, it takes two numbers off the spec sheet, and it pred

model-releasesr-ollama
30 Jul 2026
Model Releases

Mergeable Model-Side Aggregation States for Long-Context Language Models

DGX agent

arXiv:2607.26448v1 Announce Type: new Abstract: A known limitation of long-context language models is their increasingly unreliable performance in non-additive, set-based aggregation as context length

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Meta-Learned Reward Shaping for Reinforcement Learning from Human Feedback

DGX agent

arXiv:2607.26094v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) is the standard approach for aligning large language models with human preferences, but its quality

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Migrate your prompts to new models and optimize them on Amazon Bedrock

DGX agent

Amazon Bedrock Advanced Prompt Optimization optimizes your prompts for up to 5 models at once and compares original versus optimized performance across quality, latency, and cost. Migrate to a new mod

model-releasesaws-ml-blog
30 Jul 2026
Model Releases

MiniIO debuts AIStor Memory, the long-term memory AI agents need to scale safely

DGX agent

Object storage software company MiniIO Inc. says it has cracked the persistent memory problem for artificial intelligence agents with the launch of a new offering called AIStor Memory. Whereas convent

model-releasessiliconangle
30 Jul 2026
Model Releases

Mixture-of-experts for handwriting trajectory reconstruction from IMU sensors

DGX agent

arXiv:2607.26708v1 Announce Type: new Abstract: The use of digital pens for online handwriting trajectory reconstruction is a prevalent method for human-computer interaction. In this study, we focus o

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

ML2B: Benchmarking LLMs on Cross-Lingual ML Pipeline Generation

DGX agent

arXiv:2509.22768v3 Announce Type: replace Abstract: We introduce ML2B, the first benchmark for evaluating cross-lingual task comprehension in end-to-end ML pipeline generation by large language models

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Nanbeige4.2-3B: I'm not impressed

DGX agent

I've tested Nanbeige-4.2-3B. On paper, the benchmarks promise it blows away Qwen3.5-9B and Gemma4-12B. My goal was to have something very light and fast to replace Qwen3.6-35B (or finetunes thereof) f

model-releasesr-localllama
30 Jul 2026
Model Releases

Not every page in your PDF needs the same treatment. A scanned cover, a dense table, a clean text page, a figure-heavy diagram.... most pars…

DGX agent

Not every page in your PDF needs the same treatment. A scanned cover, a dense table, a clean text page, a figure-heavy diagram.... most parsing pipelines throw all of them at the same parser, forcing

model-releasesjerry-liu--x
30 Jul 2026
Model Releases

OmegaUse-OfficeVal: Benchmarking LLM Agents on Long-Horizon Office-Suite Tasks with Economic Grounding

DGX agent

arXiv:2607.27155v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly expected to assist users in completing tasks. However, existing benchmarks provide limited support

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

OmniAD: Detect and Understand Industrial Anomaly via Multimodal Reasoning

DGX agent

arXiv:2505.22039v2 Announce Type: replace Abstract: While anomaly detection has made significant progress, generating detailed analyses that incorporate industrial knowledge remains a challenge. To ad

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Online Handwriting Trajectory Reconstruction from Kinematic Sensors using Temporal Convolutional Network

DGX agent

arXiv:2607.26733v1 Announce Type: new Abstract: Handwriting with digital pens is a common way to facilitate human-computer interaction through the use of Online Handwriting (OH) trajectory reconstruct

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

OpenAI says it is cutting the price of GPT-5.6 Luna by ~80% and the price of GPT-5.6 Terra by 20% after improving the efficiency of the systems that serve them (Ina Fried/Axios)

DGX agent

Ina Fried / Axios: OpenAI says it is cutting the price of GPT-5.6 Luna by ~80% and the price of GPT-5.6 Terra by 20% after improving the efficiency of the systems that serve them — OpenAI said Thursda

model-releasestechmeme
30 Jul 2026
Model Releases

Our partners at @depthfirstlabs just released dfs-large1, a specialized model built for finding and validating real vulnerabilities in large…

DGX agent

Our partners at @depthfirstlabs just released dfs-large1, a specialized model built for finding and validating real vulnerabilities in large enterprise codebases. We helped them to scale the training

model-releasesfireworks-ai--x
30 Jul 2026
Model Releases

P.A.I. — Sleek Native Desktop AI Overlayer for Local Ollama Models 🤖⚡

DGX agent

Greetings Community! 👋 I hope everyone is doing well! I'm Tauhid — Senior EEE student from a Bangladeshi University Today I'd like to share an open-source project I’ve been developing called P.A.I. (P

model-releasesr-ollama
30 Jul 2026
Model Releases

Parameter-Free Dynamic Regret for Online Convex Optimization under Heavy-Tailed Noise

DGX agent

arXiv:2607.27073v1 Announce Type: new Abstract: We study online convex optimization (OCO) in non-stationary environments under heavy-tailed noise, where the stochastic gradient oracle admits only a fi

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

PatchDenoiser: Parameter-efficient multi-scale patch learning and fusion denoiser for Low-dose CT imaging

DGX agent

arXiv:2602.21987v3 Announce Type: replace Abstract: Low-dose CT images are essential for reducing radiation exposure in cancer screening, pediatric imaging, and longitudinal monitoring protocols, but

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Persistence Spheres: a Bi-continuous Linear Representation of Measures for Partial Optimal Transport

DGX agent

arXiv:2603.15384v2 Announce Type: replace-cross Abstract: We improve and extend persistence spheres, introduced in~ite{pegoraro2025persistence}. Persistence spheres map an integrable measure mu on the

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Phoneme- vs. Character-Level Targets and Selective State-Space Models for Intracortical Brain-to-Text

DGX agent

arXiv:2607.26751v1 Announce Type: new Abstract: State-of-the-art intracortical brain-to-text systems pair a neural-sequence phone decoder with an external language model. Two design axes remain undere

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

PolyAI launches new real-time voice conversation model to make AI-driven calls more human

DGX agent

Voice assistant and conversational AI agent developer PolyAI Ltd. today announced the release of Dialog-RSN-1, a voice dialog artificial intelligence model capable of directly perceiving and respondin

model-releasessiliconangle
30 Jul 2026
Model Releases

Position: Evaluation Scores Are Perishable Knowledge Claims

DGX agent

arXiv:2607.26191v1 Announce Type: cross Abstract: Evaluation methodologies for language models increasingly combine multiple signals, from automated metrics and LLM-as-judge ratings to human assessmen

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Post-Training at the Edge of Detectability: A Game-Theoretic Approach to Fine-Tuning

DGX agent

arXiv:2607.26358v1 Announce Type: new Abstract: Reinforcement learning (RL) fine-tuning is widely used in language model training to improve model performance on a target task while limiting drift fro

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

PowerAtlas: Towards Electricity-Computing Co-Scheduling for Power Systems

DGX agent

arXiv:2607.26710v1 Announce Type: new Abstract: The rapid growth of AI workloads is turning data centers into large-scale, volatile, yet spatiotemporally flexible grid loads, creating an urgent need f

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Progressive Multimodal Alignment for Continual Instruction Tuning

DGX agent

arXiv:2607.26947v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) rely on a projector to align visual representations with the language embedding space, making it central to cro

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Projective Graph Residualization: Variation-Allocation Frontiers for Control-Function IV

DGX agent

arXiv:2606.14636v2 Announce Type: replace Abstract: Control-function instrumental-variable estimators pass an estimated first-stage residual to an outcome model. The residual must retain the latent co

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Prosody-driven Jailbreaks in Audio LLMs: A Controlled Study and Mechanistic Analysis

DGX agent

arXiv:2607.26541v1 Announce Type: cross Abstract: Audio-capable foundation models enable end-to-end spoken interaction, but they also introduce safety risks beyond transcript content. It remains uncle

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Quick reminder of what's ok vs not ok with harnesses used for playing ARC-AGI-3: 1. Not okay: harnesses that were custom-made to solve the b…

DGX agent

Quick reminder of what's ok vs not ok with harnesses used for playing ARC-AGI-3: 1. Not okay: harnesses that were custom-made to solve the benchmark or that contain knowledge about the benchmark forma

model-releasesfrancois-chollet--x
30 Jul 2026
Model Releases

Rad-JEPA 3D: Radiology Joint-Embedding Predictive Model for 3D Computed Tomography

DGX agent

arXiv:2607.26196v1 Announce Type: new Abstract: Self-supervised pretraining is central to 3D medical image analysis, where unlabeled CT volumes are abundant but expert annotations are scarce. Yet exis

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

RAGuard: A Layered Defense Framework for Retrieval-Augmented Generation Systems Against Data Poisoning

DGX agent

arXiv:2607.26339v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems ground large language models (LLMs) in external corpora, but this reliance exposes them to corpus poisoning

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

ReCo: Reweighting GRPO Against Distributional Concentration

DGX agent

arXiv:2607.26862v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has become a standard reinforcement learning method for post-training language models. Recent work shows that

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Reeling It In: Flexible Needle Pick Up via Thread Manipulation for Autonomous Suturing

DGX agent

arXiv:2607.26337v1 Announce Type: new Abstract: Suture-needle pickup is necessary for autonomous suturing, as a needle can be unexpectedly dropped or strategically released to adjust the grasping conf

model-releasesarxiv-cs-ro
30 Jul 2026
Model Releases

Rethinking Self-Evolution: A Constrained Exploration-Exploitation Process for Mitigating Skill Overfitting

DGX agent

arXiv:2607.26643v1 Announce Type: cross Abstract: Enabling large language model (LLM) agents to accumulate and reuse experience from past interactions remains a central challenge in real-world applica

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Risk-Aware Motion Planning with Learned Trajectory Primitives and Probabilistic Safety Assessment

DGX agent

arXiv:2607.26802v1 Announce Type: new Abstract: This paper presents a radial basis function network (RBFN)-informed motion planning framework for safe and efficient urban autonomous driving. The propo

model-releasesarxiv-cs-ro
30 Jul 2026
Model Releases

Route by Kinematics, Act by Observation: Kinematics-Supervised Expert Routing in MoE-Augmented VLA

DGX agent

arXiv:2607.26807v1 Announce Type: new Abstract: While MoE augments VLA via expert specialization, router suffers from ineffective expert routing owing to the kinematic heterogeneity of actions across

model-releasesarxiv-cs-ro
30 Jul 2026
Model Releases

Same Evidence, Different Target: Decoding How Diagnostic Evidence Bears on Causal Questions from Language-Model States

DGX agent

arXiv:2607.26929v1 Announce Type: new Abstract: The same diagnostic result can support or challenge one causal claim yet fail to address another when the claims concern different populations, outcomes

model-releasesarxiv-cs-cl
30 Jul 2026
← Previous
1…6667686970…469
Next →