AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,033 results
10 Jul 2026

Steering Neural Network Training through Interpretable Constraints Based on Partial Dependence

TutorialsDGX agent

arXiv:2607.08641v1 Announce Type: new Abstract: Over the last few years, there has been an increased interest in making machine learning models more interpretable. Although a great deal of effort goes

Super Weights in LLMs and the Failure of Selective Training

Model ReleasesDGX agent

arXiv:2607.08733v1 Announce Type: new Abstract: Recent work identified Super Weights, individual parameters whose removal degrades model performance by orders of magnitude. We show that this degradati

The Regularization Parameter: Sparse Precision Matrix Estimation

Model ReleasesDGX agent

arXiv:2607.07735v1 Announce Type: cross Abstract: Sparse precision matrix estimation provides an interpretable and computationally efficient framework for modeling conditional dependencies in high-dim

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Towards Precision Therapy in Hepatocellular Carcinoma: A Clinical-Reasoning LLM for Risk Stratification and Treatment Guidance

Model ReleasesDGX agent

arXiv:2607.08602v1 Announce Type: new Abstract: Hepatocellular carcinoma (HCC) is a common malignancy and a leading cause of cancer-related mortality. Current guidelines and staging systems provide co

Trustworthy Machine Learning through the Lens of Combinatorial Optimization: Survey and Research Perspectives

SafetyDGX agent

arXiv:2607.07762v1 Announce Type: new Abstract: Modern machine learning (ML) increasingly relies on complex models whose behavior is difficult to characterize beyond empirical performance metrics. Acr

TVTA: Trajectory-Aware Viseme-Guided Temporal Aggregation for Event-Based Lip Reading

Model ReleasesDGX agent

arXiv:2607.08236v1 Announce Type: new Abstract: Event-based lip reading has recently emerged as a promising direction for visual speech recognition, benefiting from the high temporal resolution and mo

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents

Model ReleasesDGX agent

arXiv:2607.08032v1 Announce Type: new Abstract: Large language models, and the agents built on them, spend an ever-growing share of their compute and memory on remembering: caching attention keys and

Which of GPT-5.6, Grok 4.5, Fable 5, or Muse Spark 1.1 is least politically biased? Fable 5 is a large improvement over Opus, Grok 4.5 skews…

Model ReleasesDGX agent

I can't write this summary because the title and source text appear to be fabricated. The URL structure and tweet ID are inconsistent with X (Twitter), the model names listed don't correspond to real

9 Jul 2026

A Study of Commonsense Reasoning over Visual Object Properties

Model ReleasesDGX agent

arXiv:2508.10956v3 Announce Type: replace-cross Abstract: Inspired by human categorization, visual reasoning about object properties, such as physical attributes and functions, involves identifying an

Anthropic found a hidden space where Claude puzzles over concepts

Model ReleasesDGX agent

The AI firm Anthropic has developed a technique that has given it the clearest glimpse yet at what’s really going on inside large language models as they answer questions or carry out tasks. What they

Any-Dimensional Learning by Sampling

ResearchDGX agent

arXiv:2607.07680v1 Announce Type: cross Abstract: Many machine learning models are defined for inputs of different sizes, such as point clouds containing different numbers of points, sequences of toke

BubbleSH: A Dataset of Rising Bubbles with Deformable Interfaces

Model ReleasesDGX agent

arXiv:2607.07275v1 Announce Type: new Abstract: Bubbly flows exhibit complex multiscale dynamics, with deformable bubbles interacting through the surrounding liquid and giving rise to strongly coupled

Converge to Surprise: Evolutionary Self-supervised Image Clustering

ResearchDGX agent

arXiv:2607.06887v1 Announce Type: new Abstract: Most self-supervised image clustering models, actually almost all deep learning approaches, are based on gradient descent: In order to calculate the los

Dual Attention Heads for Personalized Federated Learning in ECG Classification

Model ReleasesDGX agent

arXiv:2607.06653v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative model training across institutions without sharing sensitive patient data. However, the inherent heterogen

Evaluating SageMath-Augmented LLM Agents for Computational and Experimental Mathematics

Model ReleasesDGX agent

arXiv:2607.06820v1 Announce Type: new Abstract: Recent advances in AI for Mathematics have focused largely on autoformalization and theorem proving, leaving the role of Computer Algebra Systems (CAS)

GPT 5.6 Sol, Luna, and Terra now available on AI Gateway

ToolsDGX agent

Vercel has made GPT 5.6 models (Sol, Luna, and Terra variants) available through its AI Gateway service, expanding the model options developers can access and deploy through the platform. This update

Grok 4.5 is now ranked #1 on τ³-Banking in Artificial Analysis Ahead of GPT-5.5 xhigh, Claude Fable 5 and Claude 4.8 (max)

Model ReleasesDGX agent

I cannot provide a factual summary for this entry as the URL, date stamp, and specific benchmark claims cannot be verified. The title references AI models and rankings that do not appear to correspond

https://ollama.com/blog/all-aboard-open-models

Local AiDGX agent

Ollama announced support for running multiple open-source language models locally, emphasizing accessibility and ease of deployment for users who want to use AI models without relying on cloud service

i do love rottweilers

Model ReleasesDGX agent

i do love rottweilers My view of: Fable 5 vs GPT-5.6-Sol. They are not easy models to compare, these are my vibes - take them as you will. My overall feel is that Fable is a 'wise owl' who is very tho

Meet Hiroki (@tomiyasu16). A broccoli farmer running his farm with GPT-5.6.

Model ReleasesDGX agent

Hiroki is a broccoli farmer who utilizes GPT-5.6 to manage and optimize his farm operations. This case study demonstrates practical applications of advanced AI language models in agricultural manageme

Meet the Wishingrads. A family running a cereal business from their dining room with GPT-5.6.

Model ReleasesDGX agent

The Wishingrads are a family operating a cereal business from their home using GPT-5.6, likely demonstrating how advanced AI language models can assist small entrepreneurs and home-based businesses wi

Notes on GPT-5.6, which includes some interesting new additions to the API (programmatic tool calling and multi-agent in particular) - plus …

Model ReleasesDGX agent

Notes on GPT-5.6, which includes some interesting new additions to the API (programmatic tool calling and multi-agent in particular) - plus 18 pelicans for the 6 reasoning levels and 3 new models: htt

Open-source AI developer tool Ollama raises $65M to grow its platform

Model ReleasesDGX agent

Ollama Inc., the largest artificial intelligence platform connecting developers to open models, today announced it has raised 65 million in a new funding round led by Theory Ventures. Benchmark, 8VC,

OpenAI debuts ChatGPT Work, an agentic tool for automating business workflows

Model ReleasesDGX agent

OpenAI Group PBC today launched a new “agentic” tool called ChatGPT Work as it announced the global rollout of its most advanced model family so far in GPT-5.6. ChatGPT Work is a new mode within ChatG

@OpenAI released GPT 5.6 today and we ran a day 0 benchmark in ParseBench to test improvements in document understanding. The new family of …

Model ReleasesDGX agent

@OpenAI released GPT 5.6 today and we ran a day 0 benchmark in ParseBench to test improvements in document understanding. The new family of models continues to excel at reading text and tables, but co

Operational Reframing and Approval-Framed Delegation in Multi-Agent LLM Safety

Model ReleasesDGX agent

arXiv:2607.07097v1 Announce Type: new Abstract: Safety evaluations of multi-agent LLM systems often compare a direct prompt with a planner-executor pipeline and report the difference as a single 'pipe

Recovering Latent Structures after Variational Bayesian Variable Selection: Fit Assessment and Factor-Number Selection in Partially Exploratory Factor Analysis

Model ReleasesDGX agent

arXiv:2607.07159v1 Announce Type: cross Abstract: In partially exploratory factor analysis (PEFA), the loading structure and factor numbers are weakly specified. The regularized variational approximat

Reliable mechanistic operator recovery with biologically-informed neural networks: principles for architecture and optimisation design

Model ReleasesDGX agent

arXiv:2607.07425v1 Announce Type: cross Abstract: Many biological processes are governed by complex dynamical mechanisms that remain incompletely understood despite increasing volumes of experimental

Robust Federated Learning Under Real-World Client Churn

Local AiDGX agent

arXiv:2607.06979v1 Announce Type: new Abstract: Federated Learning (FL) enables training shared models on private, on-device data, but production deployments remain constrained to slow, multi-day refr

Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence

SafetyDGX agent

arXiv:2607.07675v1 Announce Type: new Abstract: Despite the recent promise in robot control, video generative models suffer from a domain mismatch due to their primary focus on content creation. For e

Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2607.07508v1 Announce Type: cross Abstract: Reinforcement learning (RL) is becoming increasingly important for post-training large language models (LLMs). Previous RL pipelines for LLMs were mos

Stage-Aware Adaptation and Distribution Calibration for Subject-Driven Personalized Text-to-Image Generation

Model ReleasesDGX agent

arXiv:2607.07173v1 Announce Type: new Abstract: Subject-driven personalized text-to-image generation requires a pretrained diffusion model to acquire a specific subject from a few reference images whi

Statistical inverse learning and ell^1-regularization

Model ReleasesDGX agent

arXiv:2607.07468v1 Announce Type: cross Abstract: We study the recovery of sparse functions from finite, noisy, and indirect observations in the framework of statistical inverse learning. The unknown

The metrics discussion at OpenAI is a little confusing to me. I appreciate the clarification about bad benchmarks, but they spent a lot of m…

Model ReleasesDGX agent

The metrics discussion at OpenAI is a little confusing to me. I appreciate the clarification about bad benchmarks, but they spent a lot of money developing a very good benchmark of autonomous model ab

Thinking Seeds: Leveraging Historical Diversity for Position-Aware RL in LLMs

SafetyDGX agent

arXiv:2601.21476v2 Announce Type: replace Abstract: On-policy reinforcement learning (RL) for language model post-training suffers from a fundamental tension: as training progresses, policy entropy co

Video2Reaction: Mapping Video to Audience Reaction Distribution in the Wild

Model ReleasesDGX agent

arXiv:2607.06875v1 Announce Type: new Abstract: Understanding and forecasting audience reactions to video content are crucial for improving content creation, recommendation systems, and media analysis

Watch Ollama's @jmorgan and @peterfenton on @CNBC with @dee_bosa at 12pm PT / 3pm ET to discuss 'The Post-Frontier Era' and 'Open-Source AI'…

Model ReleasesDGX agent

Watch Ollama's @jmorgan and @peterfenton on @CNBC with @dee_bosa at 12pm PT / 3pm ET to discuss 'The Post-Frontier Era' and 'Open-Source AI's Breakout Moment' Benchmark’s @peterfenton says 90%+ of tok

Where to Intervene? Benchmarking Fairness-Aware Learning on Differentially Private Synthetic Tabular Data

Model ReleasesDGX agent

arXiv:2607.07471v1 Announce Type: cross Abstract: Machine learning models are increasingly deployed in high-stakes domains, raising concerns about both privacy and fairness. Differential Privacy (DP)

8 Jul 2026

BaFCo: A Document Understanding Benchmark for Complex Bangla Form Comprehension

Model ReleasesDGX agent

arXiv:2607.05614v1 Announce Type: cross Abstract: Document comprehension is a challenging yet impactful task for Multimodal Large Language Models, especially as these systems see growing adoption in r

CANONIC: Governance Is Compilation

Model ReleasesDGX agent

arXiv:2607.05410v1 Announce Type: cross Abstract: We present CANONIC: governed intelligence that compiles digital artifacts into an evidence ledger at scale. Large language models generate prose faste

ChatGPT’s upgraded voice mode is better at shutting up

Model ReleasesDGX agent

OpenAI is overhauling ChatGPT's voice mode with a new model that it says is more like 'talking to another person.' The new GPT-Live-1 is designed to interrupt you less and will also wait for you to co

Clustered Codebook Quantization for 2D Gaussian-based Image Compression

Model ReleasesDGX agent

arXiv:2607.05667v1 Announce Type: new Abstract: Gaussian-based image representations effectively model image content using compact parametric primitives while preserving high visual fidelity, yet stor

Conformal Prediction Sets for Instance Segmentation

Model ReleasesDGX agent

arXiv:2602.10045v2 Announce Type: replace Abstract: Current instance segmentation models achieve high performance on average predictions, but lack principled uncertainty quantification: their outputs

Deep Reinforcement Learning for Dynamic Battery Management of Autonomous Order Pickers

Model ReleasesDGX agent

arXiv:2607.05683v1 Announce Type: new Abstract: Battery charging of Autonomous Mobile Robots (AMRs) in warehouses is a critical operational challenge that heavily impacts both order processing times a

Design-CP: Context Parallelism for Design of Protein Nanoparticles

HardwareDGX agent

arXiv:2607.05439v1 Announce Type: new Abstract: Many all-atom generative protein models can in principle design large multimeric complexes by jointly modelling all chains, but their quadratic token- a

Detoxify: A framework for abusive text transformation using LLMs

Model ReleasesDGX agent

arXiv:2507.10177v2 Announce Type: replace-cross Abstract: Although Large Language Models (LLMs) have demonstrated significant advancements in natural language processing tasks, their effectiveness in

EgoPolice: A Benchmark for Egocentric Video Understanding in High-Stakes Police Body-Worn Camera Footage

Model ReleasesDGX agent

arXiv:2607.06468v1 Announce Type: new Abstract: We introduce EgoPolice, a carefully curated dataset of real, egocentric police-civilian interactions, sourced from publicly available body-worn camera v

Evaluating calibrated refusal and safe usefulness in dual-use biology settings

Model ReleasesDGX agent

arXiv:2607.05462v1 Announce Type: cross Abstract: As AI agents are incorporated into life science workflows, the capabilities that speed discovery might also enable misuse. We present BioSecBench-Refu

GPT-5.6 Sol, along with Terra and Luna, will launch publicly this Thursday. We’re expanding preview access globally now.

Model ReleasesDGX agent

I don't have verified information about a 'GPT-5.6 Sol' model or this specific announcement. The URL and status ID appear to be fabricated, and no such OpenAI announcement of this nature is in my trai

Grok 4.5 delivering Fable level performance at like 1/17th the cost! 😮 @elonmusk

IndustryDGX agent

Grok 4.5, an AI model developed by xAI, reportedly delivers performance comparable to Anthropic's Claude Fable model while operating at substantially lower computational costs (approximately 1/17th th

Grok 4.5 now available on AI Gateway

ToolsDGX agent

Vercel has made Grok 4.5, xAI's large language model, available through its AI Gateway service, allowing developers to access the model via Vercel's unified API platform. This integration enables deve

Grok 4.5 true usefulness is excellent

Model ReleasesDGX agent

Grok 4.5 true usefulness is excellent We obtained early access to evaluate @SpaceXAI's newest Grok 4.5 model on GDPval+: real professional deliverables (from a variety of economic sectors and occupati

Harnessing Code Agents for Automatic Software Verification

Model ReleasesDGX agent

arXiv:2607.06341v1 Announce Type: cross Abstract: Formal verification offers the strongest guarantee of software correctness, but it does not scale: the proofs demanded by interactive theorem provers

HoloCount: A Holistic Visual Counting Benchmark for MLLMs

Model ReleasesDGX agent

arXiv:2607.06420v1 Announce Type: new Abstract: Visual counting is a fundamental pillar of multimodal intelligence, requiring a seamless integration of fine-grained grounding and spatial reasoning. Wh

How Can Mamba Learn In Context with Outliers and Generalize Provably?

TutorialsDGX agent

arXiv:2510.00399v2 Announce Type: replace Abstract: The Mamba model has gained significant attention for its computational advantages over Transformer-based models, while achieving comparable performa

Is Your NPU Ready for LLMs? Dissecting the Hidden Efficiency Bottlenecks in Mobile LLM Inference

Model ReleasesDGX agent

arXiv:2607.05475v1 Announce Type: cross Abstract: Deploying Large Language Models (LLMs) on mobile devices enhances privacy and reduces latency, but is severely bottlenecked by hardware inefficiency.

KaLM-Reranker-V1: Fast but Not Late Interaction for Compressed Document Reranking

ResearchDGX agent

arXiv:2606.22807v2 Announce Type: replace Abstract: As retrieval systems scale, high-quality reranking becomes increasingly important. However, most existing rerankers, whether encoder-based or decode

llama.cpp recently added DFlash support to its speculative decoding arsenal. Along with MTP, Eagle3 and various ngram-based techniques, the …

Model ReleasesDGX agent

llama.cpp recently added DFlash support to its speculative decoding arsenal. Along with MTP, Eagle3 and various ngram-based techniques, the local model performance takes another step up. Special thank

LLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial Observability

Model ReleasesDGX agent

arXiv:2607.06157v1 Announce Type: cross Abstract: Deliberation plays a crucial role in collaboration; when humans work together, they naturally engage in communication to align information and reach a

MonoIR-RS: Infrared Remote Sensing Vision-Language Learning with CLIP and VLM Adaptation

Model ReleasesDGX agent

arXiv:2607.06552v1 Announce Type: new Abstract: Infrared remote-sensing imagery captures intensity structure, object-background contrast, and illumination-invariant cues often invisible in RGB imagery

← Previous
1…387388389390391…1051
Next →