AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,929 results
Research

From Classical Machine Learning to Tabular Foundation Models: An Empirical Investigation of Robustness and Scalability Under Class Imbalance in Emergency and Critical Care

DGX agent

arXiv:2512.21602v2 Announce Type: replace-cross Abstract: Millions of patients pass through emergency departments and intensive care units each year, where clinicians must make high-stakes decisions u

researcharxiv-cs-cv
10 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

I strongly suspect that Claude Mythos is a looped language model, as described in the paper 'Scaling Latent Reasoning via Looped Language Mo…

DGX agent

I strongly suspect that Claude Mythos is a looped language model, as described in the paper 'Scaling Latent Reasoning via Looped Language Models' from ByteDance The authors of that paper called out gr

model-releasesjeremy-howard--x
10 Apr 2026
Model Releases

Illocutionary Explanation Planning for Source-Faithful Explanations in Retrieval-Augmented Language Models

DGX agent

arXiv:2604.06211v1 Announce Type: cross Abstract: Natural language explanations produced by large language models (LLMs) are often persuasive, but not necessarily scrutable: users cannot easily verify

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

MM-MoralBench: A MultiModal Moral Evaluation Benchmark for Large Vision-Language Models

DGX agent

arXiv:2412.20718v2 Announce Type: replace Abstract: The rapid integration of Large Vision-Language Models (LVLMs) into critical domains necessitates comprehensive moral evaluation to ensure their alig

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

TalkLoRA: Communication-Aware Mixture of Low-Rank Adaptation for Large Language Models

DGX agent

arXiv:2604.06291v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) enables parameter-efficient fine-tuning of Large Language Models (LLMs), and recent Mixture-of-Experts (MoE) extensions fur

model-releasesarxiv-cs-ai
10 Apr 2026
Agents

TOOLCAD: Exploring Tool-Using Large Language Models in Text-to-CAD Generation with Reinforcement Learning

DGX agent

arXiv:2604.07960v1 Announce Type: cross Abstract: Computer-Aided Design (CAD) is an expert-level task that relies on long-horizon reasoning and coherent modeling actions. Large Language Models (LLMs)

agentsarxiv-cs-cl
10 Apr 2026
Local Ai

Use the Same Model Across Ollama, LM Studio, Jan, and your Favorite Local AI Apps

DGX agent

Local AI tools such as Ollama, LM Studio, and Jan all rely on the same underlying inference engine (llama.cpp) and support compatible model formats (primarily GGUF), meaning a single downloaded mod...

local-air-ollama
9 Apr 2026
Tools

[AINews] Meta Superintelligence Labs announces Muse Spark, first frontier model on their completely new stack

DGX agent

Muse Spark is the first model from Meta's Superintelligence Labs division and the debut entry in the new Muse model family, representing a 'ground-up overhaul' of the company's AI efforts. It is ...

toolslatent-space
8 Apr 2026
Industry

Only a few hundred Tesla Model S & X cars left in inventory. Order now if you want one.

DGX agent

In April 2026, Elon Musk announced via X that only a few hundred Tesla Model S and Model X units remained in inventory, urging prospective buyers to order immediately. After more than a decade on ...

industryelon-musk--x
8 Apr 2026
Model Releases

I spent the night testing open-source coding models against Claude Opus in production. Same codebase. Same tasks. Real API calls, real file …

DGX agent

I spent the night testing open-source coding models against Claude Opus in production. Same codebase. Same tasks. Real API calls, real file edits, real bugs. Tested: Arcee Trinity-Large-Thinking, http

model-releaseszhipu-ai--x
7 Apr 2026
Safety

BEST-KAG: Enhancing Question Answering of Building Engineering Standards with Multimodal Knowledge Graph Modeling and Large Language Model

DGX agent

arXiv:2608.11244v1 Announce Type: new Abstract: Construction standards are critical for building safety and sustainability. Existing standard application workflows rely on keyword-based document retri

safetyarxiv-cs-ai
13 Aug 2026
Model Releases

Do You See What You Draw? A Semantic Closed-Loop Framework for Holistic Evaluation of Unified Multimodal Models

DGX agent

arXiv:2608.11907v1 Announce Type: cross Abstract: As Large Vision-Language Models increasingly aim to integrate visual generation and understanding within a single parameter space, evaluating such str

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Locating and Controlling Implicit Personalization in Large Language Models

DGX agent

arXiv:2608.11735v1 Announce Type: cross Abstract: Large language models (LLMs) often shift their outputs in response to implicit demographic cues even when users never state a demographic identity. Pr

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Our most intelligent workhorse model yet for coding and agents has arrived ⚡ Meet Gemini 3.7 Flash. — Crush that seemingly endless to-do lis…

DGX agent

Our most intelligent workhorse model yet for coding and agents has arrived ⚡ Meet Gemini 3.7 Flash. — Crush that seemingly endless to-do list. Gemini Spark in the @geminiapp now uses 3.7 Flash. The ne

model-releasesgoogle-ai--x
13 Aug 2026
Safety

Post-Training with Policy Gradients: Optimality and the Base Model Barrier

DGX agent

arXiv:2603.06957v2 Announce Type: replace-cross Abstract: We study post-training linear autoregressive models with outcome and process rewards. Given a context oldsymbol{x}, the model must predict the

safetyarxiv-cs-ai
13 Aug 2026
Model Releases

Reliable Inference in Edge-Cloud Model Cascades via Conformal Alignment

DGX agent

arXiv:2510.17543v3 Announce Type: replace Abstract: Edge intelligence enables low-latency inference via compact on-device models, but assuring reliability remains challenging. We study edge-cloud casc

model-releasesarxiv-cs-lg
13 Aug 2026
Model Releases

Repurposing RGB-based Foundation Model for Depth Estimation on Thermal Images Using Hierarchical Supervision

DGX agent

arXiv:2608.11564v1 Announce Type: new Abstract: Depth estimation from thermal images is highly valuable for robotic applications in adverse conditions, such as nighttime and rainy weather. Recent stud

model-releasesarxiv-cs-cv
13 Aug 2026
Model Releases

RoboHarness: A Memory-Augmented Policy Harness for Vision-Language-Action Model Robustness via In-Context Adaptation

DGX agent

arXiv:2603.24060v3 Announce Type: replace Abstract: Despite the promise of Vision-Language-Action (VLA) models as generalist robotic controllers, their robustness against perceptual noise and environm

model-releasesarxiv-cs-ro
13 Aug 2026
Applications

Robust Ambiguity Detection (RAD) From Model- and Feature-Space Consistency

DGX agent

arXiv:2608.11541v1 Announce Type: new Abstract: Machine learning models should be robust, in the sense of remaining predictively consistent under permissible variations. A model's predictions should i

applicationsarxiv-cs-lg
13 Aug 2026
Research

Task- and dataset-specific information in protein language models

DGX agent

arXiv:2608.12090v1 Announce Type: new Abstract: Protein language models (PLMs) have transferred the latest advances from natural language processing to computational biology. These models, trained on

researcharxiv-cs-lg
13 Aug 2026
Safety

Do AI weather models miss extremes?

DGX agent

arXiv:2608.09972v1 Announce Type: cross Abstract: First-generation AI weather models are often reported to underperform at extremes, mostly in reanalysis-based evaluations of deterministic regression

safetyarxiv-cs-ai
12 Aug 2026
Research

DoseBridge: Denoising Diffusion Bridge Model for Dose Prediction in Lung Intensity-Modulated Proton Therapy

DGX agent

arXiv:2608.10173v1 Announce Type: new Abstract: Most radiotherapy dose-prediction models use only CT images and anatomical structures, although intensity-modulated proton therapy (IMPT) dose also depe

researcharxiv-cs-cv
12 Aug 2026
Safety

Dreamer-SAC: Off-Policy Learning in Latent World Models for Sample-Efficient Autonomous Driving

DGX agent

arXiv:2608.10386v1 Announce Type: new Abstract: Sample-efficient reinforcement learning for autonomous driving is often limited by the trade-off between data efficiency and model bias. While world mod

safetyarxiv-cs-lg
12 Aug 2026
Model Releases

give 4.6 a try and let us know how it goes. your feedback is a big part of why the model gets better with each iteration.

DGX agent

give 4.6 a try and let us know how it goes. your feedback is a big part of why the model gets better with each iteration. SpaceXAI's Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index, j

model-releaseselon-musk--x
12 Aug 2026
Research

MoE Proxy Models for Low-Cost Failure Reproduction and Diagnosis in LLM RL Post-Training

DGX agent

arXiv:2608.10823v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training of large language models (LLMs) is computationally intensive and involves complex system pipelines with substa

researcharxiv-cs-lg
12 Aug 2026
Model Releases

ChronoState: Hidden Elapsed-Time Conditioning for Temporal-State Action Selection in Frozen-Backbone Language Models

DGX agent

arXiv:2608.09124v1 Announce Type: new Abstract: Temporal decisions in language-model systems often depend on both symbolic task state and elapsed wall-clock time, such as cache expiration, job complet

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

ComplexityWorld: Benchmarking Vision-Language Models on Verifiable Visual Decision Making

DGX agent

arXiv:2608.07584v1 Announce Type: new Abstract: Vision-language models (VLMs) have made rapid progress in visual perception and increasingly support real-world tasks that depend on images. Many such t

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

EmoS: A Theory-Grounded Framework for Evaluating and Aligning Emotional Intelligence in Spoken Language Models

DGX agent

arXiv:2608.09189v1 Announce Type: new Abstract: Despite significant advances in instruction-following and auditory comprehension, the evaluation of Emotional Intelligence (EI) in Spoken Language Model

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

FoMoH: A clinically meaningful foundation model evaluation for structured electronic health records

DGX agent

arXiv:2505.16941v4 Announce Type: replace-cross Abstract: Foundation models (FMs) promise to address core limitations of traditional supervised machine learning: (i) reliance on large amounts of label

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

In-Loop Model Adaptation with Coupled Latent-Noise Guidance for High-Fidelity Subject-Driven Text-to-Image Generation

DGX agent

arXiv:2608.09244v1 Announce Type: new Abstract: Text-to-image diffusion models have achieved remarkable success in generating high-quality images from a given text prompt. Subject-driven generation ai

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

MathShikkha: A Controlled Study of Answer-Only and Chain-of-Thought Supervision for Bangla Mathematical Reasoning in Small Language Models

DGX agent

arXiv:2608.08503v1 Announce Type: new Abstract: Mathematical reasoning remains challenging in low-resource languages such as Bangla. We study whether teacher-generated Bangla Chain-of-Thought (CoT) su

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MedPixel: A Unified Pixel-Language Model for Medical Reasoning and Segmentation

DGX agent

arXiv:2608.09818v1 Announce Type: cross Abstract: Reliable medical image understanding requires models to connect clinical language and visual reasoning with pixel-level grounding. Yet medical vision-

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

OWN YOUR INTELLIGENCE Last year, building on open-weight models was primarily a cost rationalization exercise. Slightly worse performance fo…

DGX agent

OWN YOUR INTELLIGENCE Last year, building on open-weight models was primarily a cost rationalization exercise. Slightly worse performance for a much cheaper price. Now, it is increasingly an existenti

safetysonya-huang--x
11 Aug 2026
Research

Quantization Degradation in Large Language Models: A Signal-Noise Perspective

DGX agent

arXiv:2608.08188v1 Announce Type: new Abstract: Post-training quantization reduces the deployment cost of large language models, yet how severely a quantized model degrades is not determined by bit-wi

researcharxiv-cs-ai
11 Aug 2026
Model Releases

SDDBMs: Soft Denoising Diffusion Bridge Models

DGX agent

arXiv:2608.08594v1 Announce Type: new Abstract: Diffusion bridge models leverage Doob's (h)-transform to construct stochastic transports between arbitrary endpoint distributions, and have shown strong

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Sekai2: From World Exploration to Interactive World Modeling

DGX agent

arXiv:2608.09449v1 Announce Type: new Abstract: Video world models must capture how scenes evolve over time and across viewpoints. Training them for long-horizon generation and camera control therefor

model-releasesarxiv-cs-cv
11 Aug 2026
Industry

Using the different models in different industries - your experience?

DGX agent

Hi all, I'm seeing so many conversations from people discussing how they're 'using the models wrong' and 'don't use Sol max as high is enough for you', yet all the conversations lack the nuance of wha

industryr-chatgpt
11 Aug 2026
Hardware

verdi: retrieval is not transfer for continual world model optimization

DGX agent

arXiv:2608.09537v1 Announce Type: new Abstract: Foundation world models have made remarkable progress in planning, simulation, and embodied intelligence. However, optimizing a pretrained world model t

hardwarearxiv-cs-ai
11 Aug 2026
Safety

World Tokens: Enhancing Embodied Policies with Training-Time World Modeling

DGX agent

arXiv:2608.09730v1 Announce Type: new Abstract: Vision-language-action (VLA) models are a widely adopted paradigm for embodied policies. They excel at efficient closed-loop control but do not explicit

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

WuYuEval: A Multi-Level Benchmark for Large Language Models in Solid Waste Management

DGX agent

arXiv:2608.07529v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as technical assistants, but their competence in solid waste management (SWM) remains difficult to

model-releasesarxiv-cs-ai
11 Aug 2026
Research

ZetaGPT: A Reference Implementation of Positional--Encoding--Free State--Space--Attention Language Models

DGX agent

arXiv:2608.09432v1 Announce Type: cross Abstract: Transformer-based language models rely on self-attention, whose computation is permutation-equivariant and therefore lacks an intrinsic mechanism for

researcharxiv-cs-ai
11 Aug 2026
Research

AfriNLLB: Efficient Translation Models for African Languages

DGX agent

arXiv:2602.09373v2 Announce Type: replace Abstract: In this work, we present AfriNLLB, a series of lightweight models for efficient translation from and into African languages. AfriNLLB supports 15 la

researcharxiv-cs-cl
10 Aug 2026
Model Releases

Ask-E: An Environment for Calibrated Question Generation

DGX agent

arXiv:2608.06933v1 Announce Type: cross Abstract: Today, we improve models by training and evaluating them on problems at the frontier of their abilities. Creating such problems is itself a demanding

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Beyond the Black Box: Interpretable Models of Human Randomisation Failures

DGX agent

arXiv:2608.07220v1 Announce Type: new Abstract: Mixed strategy equilibrium predicts i.i.d play: past actions should not help predict future decisions. Human players, however, systematically depart fro

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Control-Anchored Residual Flow Matching Conditioned on Gene Geometry for Virtual Cell Perturbation Modeling

DGX agent

arXiv:2608.06824v1 Announce Type: cross Abstract: A central task in virtual cell modeling is predicting single-cell transcriptional responses to unseen genetic perturbations and drug combinations, and

model-releasesarxiv-cs-ai
10 Aug 2026
Research

Evaluating Useful Surrogate Models for Configuration Tuning Beyond Accuracy: A Fitness Landscape Analysis Perspective

DGX agent

arXiv:2509.21945v2 Announce Type: replace-cross Abstract: To efficiently tune configuration for better software system performance (e.g., latency) at the deployment and maintenance stage, many tuners

researcharxiv-cs-ai
10 Aug 2026
Safety

Explore or Converge? Stage-Guided Per-Step Optimization for Diffusion Models

DGX agent

arXiv:2608.06768v1 Announce Type: new Abstract: Diffusion models have strong generative capabilities. However, their maximum likelihood training objective only focuses on reconstructing the data distr

safetyarxiv-cs-cv
10 Aug 2026
Model Releases

Index SLM Technical Report

DGX agent

arXiv:2607.09885v2 Announce Type: replace Abstract: We present Index-1.9B, a series of open small language models developed at Bilibili. The series comprises four models: Index-1.9B-Base, a foundation

model-releasesarxiv-cs-cl
10 Aug 2026
← Previous
1…5556575859…1249
Next →