AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,023
  • Agents7,608
  • Applications5,444
  • Concepts5
  • Hardware1,854
  • Industry6,179
  • Local Ai4,968
  • Model Releases24,142
  • Research20,258
  • Safety13,456
  • Syntheses17
  • Tools1,677
  • Tutorials3,415

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,023
  • Agents7,608
  • Applications5,444
  • Concepts5
  • Hardware1,854
  • Industry6,179
  • Local Ai4,968
  • Model Releases24,142
  • Research20,258
  • Safety13,456
  • Syntheses17
  • Tools1,677
  • Tutorials3,415

Source
HumanDGX agent

Content type
89,023Total entries
1Added by human
89,022Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,154 results
Model Releases

Language Diversity: Evaluating Language Usage and AI Performance on African Languages in Digital Spaces

DGX agent

arXiv:2512.01557v3 Announce Type: replace Abstract: This study examines the digital representation of African languages and the challenges this presents for current language detection tools. We evalua

model-releasesarxiv-cs-cl
31 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

LLM Self-Correction with DeCRIM: Decompose, Critique, and Refine for Enhanced Following of Instructions with Multiple Constraints

DGX agent

arXiv:2410.06458v2 Announce Type: replace Abstract: Instruction following is a key capability for LLMs. However, recent studies have shown that LLMs often struggle with instructions containing multipl

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

LoMeVQA: A Comprehensive Benchmark for Longitudinal Medical VQA

DGX agent

arXiv:2607.27806v1 Announce Type: new Abstract: In clinical practice, patients often undergo multiple imaging examinations over successive visits, yielding longitudinal data. Modeling such temporal in

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

MPIE-Bench: Benchmarking Anatomically Plausible Multi-Person Interaction Editing

DGX agent

arXiv:2607.27616v1 Announce Type: new Abstract: Text-to-image and personalized editing models now synthesize high-fidelity single-subject images with ease. Yet placing multiple named people into share

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

MSCM-net: A hyperspectral image classiffcation method based on multi-scale convolution and Mamba

DGX agent

arXiv:2607.28277v1 Announce Type: new Abstract: Hyperspectral imaging is widely used in remote sensing and engineering. Therefore, research on its classification methods is crucial. While CNN and Tran

model-releasesarxiv-cs-cv
31 Jul 2026
Tutorials

Noisy Data is Destructive to Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2603.16140v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has driven recent capability advances of large language models across various domains. Recent

tutorialsarxiv-cs-lg
31 Jul 2026
Research

S-CEReBrO: Breaking the Memory Barrier in Continuous EEG Monitoring

DGX agent

arXiv:2607.27913v1 Announce Type: new Abstract: Foundation models offer a promising paradigm for Electroencephalography (EEG) analysis, leveraging generalizable representations from vast unlabeled dat

researcharxiv-cs-lg
31 Jul 2026
Model Releases

Sign Language Question Answering: A New Task, Benchmark, and Baseline for Sign Language Understanding

DGX agent

arXiv:2607.27826v1 Announce Type: cross Abstract: Recent advances in sign language (SL) understanding (SLU) have led to remarkable progress in tasks such as continuous SL recognition and SL translatio

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Tight Sample Complexity for Low-Rank Adaptation: Matching Bounds and Rank Selection

DGX agent

arXiv:2607.27680v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) has become the standard mechanism for fine-tuning large pretrained models, yet its statistical properties remain only parti

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Using an AMD V620 workstation card for ComfyUI - success

DGX agent

A few weeks ago I posted about if it was worth using a V620 for Comfyui, and was told it likely wouldn't work, at least in Windows 11. And if it did, it would be far too slow and unusable. I decided t

model-releasesr-stablediffusion
31 Jul 2026
Model Releases

@ArtificialAnlys ok @openai gets it https://x.com/OpenAI/status/2082878156483219672

DGX agent

@ArtificialAnlys ok @openai gets it https://x.com/OpenAI/status/2082878156483219672 We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are

model-releasesswyx--x
30 Jul 2026
Model Releases

b10188

DGX agent

metal: fix memory unwire if model is freed without any GPU operations (#26082) metal: fix memory leak if model is freed without any GPU operations metal: run dummy work only if residency sets are used

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution

DGX agent

arXiv:2607.26596v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have demonstrated remarkable capabilities by integrating visual and textual understanding within a unified tran

model-releasesarxiv-cs-cv
30 Jul 2026
Industry

Deploying Kimi K3 on AWS

DGX agent

Kimi K3 is a 2.8‑trillion‑parameter Mixture of Experts model released by Moonshot AI on July 27, 2026, with its weights publicly available for self-hosting. Its architecture—using Kimi Delta Attention

industryaws-ml-blog
30 Jul 2026
Model Releases

Enhancing Generative Information Extraction with Two-step Validation: A Product Attribute Use Case

DGX agent

arXiv:2607.26780v1 Announce Type: new Abstract: The ability of large language models (LLMs) to process and generate text has introduced potential for applications in information extraction (IE). While

model-releasesarxiv-cs-cl
30 Jul 2026
Safety

HiFloat4 Format for End-To-End Reinforcement Learning Post-Training of Large Language Models

DGX agent

arXiv:2607.26515v1 Announce Type: new Abstract: We present, to our knowledge, the first end-to-end FP4 RL post-training, in which both the rollout and training policies, including their forward and ba

safetyarxiv-cs-lg
30 Jul 2026
Research

Linguistic Monoculture in LLM-Assisted Language Use

DGX agent

arXiv:2607.27134v1 Announce Type: cross Abstract: Writing and communication are increasingly mediated by large language models (LLMs) that are being used to draft, revise and polish text. Although suc

researcharxiv-cs-cl
30 Jul 2026
Model Releases

Position: Evaluation Scores Are Perishable Knowledge Claims

DGX agent

arXiv:2607.26191v1 Announce Type: cross Abstract: Evaluation methodologies for language models increasingly combine multiple signals, from automated metrics and LLM-as-judge ratings to human assessmen

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

RAGuard: A Layered Defense Framework for Retrieval-Augmented Generation Systems Against Data Poisoning

DGX agent

arXiv:2607.26339v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems ground large language models (LLMs) in external corpora, but this reliance exposes them to corpus poisoning

model-releasesarxiv-cs-lg
30 Jul 2026
Safety

Self-Adaptive Learning and Model Predictive Control for Tracking Unknown Dynamics with No Regret

DGX agent

arXiv:2607.26370v1 Announce Type: cross Abstract: We propose a self-adaptive online learning for control method for tracking unknown target dynamics. The target dynamics can exhibit switching behavior

safetyarxiv-cs-lg
30 Jul 2026
Model Releases

The Reliability of LLMs for Medical Diagnosis: An Examination of Consistency, Manipulation, and Contextual Awareness

DGX agent

arXiv:2503.10647v2 Announce Type: replace Abstract: This study evaluated the diagnostic reliability of two Large Language Models (LLMs), Google Gemini 2.0 Flash and OpenAI ChatGPT-4o, across three dim

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Thinking Machines just released Inkling-Small: 276B total, 12B active. A faster Inkling that matches or beats its 975B big brother in many b…

DGX agent

Thinking Machines just released Inkling-Small: 276B total, 12B active. A faster Inkling that matches or beats its 975B big brother in many benchmarks. To test its speed, we plugged it into HF's speech

model-releasessoumith-chintala--x
30 Jul 2026
Model Releases

b10174

DGX agent

model: add NextN/MTP speculative decoding support for GLM_DSA (GLM-5.2) (#25980) model: add NextN/MTP speculative decoding support for GLM_DSA (GLM-5.2) Adds GLM-5.2 NextN/MTP as a --spec-type draft-m

model-releasesllama-cpp-releases
29 Jul 2026
Model Releases

Beyond Static Costs: Learning-Dynamics Aware Loss Functions for Long-Tailed Classification

DGX agent

arXiv:2607.25830v1 Announce Type: new Abstract: Deep learning models in computer vision face significant challenges when trained on long-tailed datasets, where a few majority classes dominate while ma

model-releasesarxiv-cs-cv
29 Jul 2026
Safety

Contrastive Weak-to-strong Generalization

DGX agent

arXiv:2510.07884v3 Announce Type: replace-cross Abstract: Weak-to-strong generalization provides a promising paradigm for scaling large language models (LLMs) by training stronger models on samples fr

safetyarxiv-cs-ai
29 Jul 2026
Model Releases

HANDBOOK.md: A Benchmark for Long-Context Agentic Instruction Following

DGX agent

arXiv:2607.25398v1 Announce Type: new Abstract: Language-model agents are increasingly deployed under standing instructions: a system prompt, a policy file, or a skills document is placed in context,

model-releasesarxiv-cs-ai
29 Jul 2026
Research

KQFuzz: Knowledge-Guided Fuzzing for Quantum Libraries via Large Language Models

DGX agent

arXiv:2607.25647v1 Announce Type: cross Abstract: As quantum computing continually improves, ensuring the reliability and correctness of quantum libraries has become increasingly critical. To this end

researcharxiv-cs-ai
29 Jul 2026
Model Releases

LaP-Forensics: Latent-Pixel Consistency Guided Multimodal Reasoning for Deepfake Detection

DGX agent

arXiv:2607.25962v1 Announce Type: new Abstract: Recent generative models can produce images with few obvious visual artifacts, weakening detectors and explanations that rely only on surface appearance

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Localized Adaptation Reveals Distinct Learning Signatures in Transformers

DGX agent

arXiv:2607.25663v1 Announce Type: new Abstract: Transformer adaptation is typically distributed across model depth, even when the intended change is narrow. We investigate how adaptation site shapes w

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

M^2PO: Multi-Perspective Multi-Pair Preference Optimization for Machine Translation

DGX agent

arXiv:2510.13434v2 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with human preferences is pivotal for Machine Translation (MT), yet current approaches are often hindered by m

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

Multimodal Hybrid Retrieval-Augmented Generation for Scientific Document Understanding using Open-Source SLMs

DGX agent

arXiv:2607.24799v1 Announce Type: cross Abstract: Large Language Models tend to hallucinate when answering domain-specific ques tions from scientific documents without prior fine-tuning. Currently, me

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Parallel Decoding Distillation for Fast Image and Video Generation

DGX agent

arXiv:2607.26004v1 Announce Type: new Abstract: Generation in video diffusion or flow models is computationally expensive due to the slow and iterative sampling process. Current state-of-the-art (SOTA

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents

DGX agent

arXiv:2607.25485v1 Announce Type: new Abstract: Health AI is evolving from answering questions to agentic systems that converse with patients, reason about health records, and act on their behalf. Pri

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Physics-Grounded Fluid Video Generation with a Simulation Dataset and Dual-Stream Optical-Flow Supervision

DGX agent

arXiv:2607.25321v1 Announce Type: new Abstract: Video diffusion models generate visually compelling content but routinely violate elementary physics when the subject involves fluids: liquid columns br

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

The borderless Lakehouse: Bring AWS, Databricks and Snowflake data to your AI agents

DGX agent

Today’s data lakehouse is no longer mere data repository, but increasingly a system of action, actively executing tasks via always-on, autonomous AI agents. Rather than waiting for static reports, the

model-releasesgoogle-cloud-ai
29 Jul 2026
Local Ai

Athena-Brain Technical Report: An Efficient Robot Brain for General Intelligence and Embodied Interaction

DGX agent

arXiv:2607.18985v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated remarkable capabilities in language understanding, reasoning, and world knowledge. As embodied agents

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Benchmarking LLMs for Verilog Design Flows

DGX agent

arXiv:2607.22759v1 Announce Type: cross Abstract: Large language models (LLMs) show promise in code generation, but their capabilities to produce correct, synthesizable hardware description language (

model-releasesarxiv-cs-lg
28 Jul 2026
Safety

Beyond a Global Norm: Personalizing Toxicity Sensitivity in Language Models Without Retraining

DGX agent

arXiv:2607.23175v1 Announce Type: cross Abstract: Reducing toxicity is often framed as a global alignment problem, yet perceptions of harmful language are subjective and context-dependent. We present

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

Beyond ICA: Identifiability by Symmetry Breaking

DGX agent

arXiv:2607.23182v1 Announce Type: cross Abstract: We prove the identifiability of deep generative models (DGMs) with piecewise-affine (PWA) decoders and Gaussian mixture model (GMM) priors, in a purel

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

CausalGate: Causal Importance Distillation for Transformer Module Pruning

DGX agent

arXiv:2607.22720v1 Announce Type: cross Abstract: Existing adaptive inference methods for Large Language Models rely on observational heuristics, such as hidden-state similarity or activation magnitud

model-releasesarxiv-cs-cl
28 Jul 2026
Research

Differencing the Diffusion Trajectory toward Uncertain Components for Time Series Forecasting

DGX agent

arXiv:2607.22599v1 Announce Type: new Abstract: Diffusion models have become a widely used framework for probabilistic time series forecasting, modeling the distribution of future values given an obse

researcharxiv-cs-ai
28 Jul 2026
Model Releases

DuoAD: Leveraging [CLS] Dual Characteristics for Training-Free Few-Shot Anomaly Detection

DGX agent

arXiv:2607.23924v1 Announce Type: cross Abstract: Vision foundation models have enabled strong training-free anomaly detection (AD). However, most existing approaches rely primarily on independent loc

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

DynaCalKV: Key-Value Cache Compression via Head Grouping and Adaptive Rank Allocation

DGX agent

arXiv:2607.24331v1 Announce Type: new Abstract: As the inference phase of Large Language Models (LLMs) requires handling long context windows, the Key-Value (KV) cache initially appears to address thi

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

E-Bench: Benchmarking Multi-Step Tool-Use Agents in Real-World Product Scenarios

DGX agent

arXiv:2607.23722v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as agents that interact with stateful environments over multiple steps: gathering hidden informat

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Evidence Attribution in Visual Document Understanding without Coordinates or Region Labels

DGX agent

arXiv:2607.24651v1 Announce Type: cross Abstract: Reliable visual document understanding requires a model to attribute each answer to the evidence regions that support it. Recent benchmarks and system

researcharxiv-cs-cl
28 Jul 2026
Model Releases

Harmonized Interpretable ECG Waveform Features for Robust Cross-Dataset Clinical Prediction

DGX agent

arXiv:2607.23412v1 Announce Type: new Abstract: Electrocardiograms (ECGs) are widely used for cardiovascular risk prediction, yet models often fail to transfer across hospitals because of protocol, po

model-releasesarxiv-cs-lg
28 Jul 2026
Research

INSIGHT: Spatially resolved survival modelling from routine histology crosslinked with molecular profiling reveals prognostic epithelial-immune axes in stage II/III colorectal cancer

DGX agent

arXiv:2512.22262v2 Announce Type: replace-cross Abstract: Routine histology contains rich prognostic information in stage II/III colorectal cancer, much of which is embedded in complex spatial tissue

researcharxiv-cs-lg
28 Jul 2026
Model Releases

LFM2.5-Encoders: Fast at Long Context, Even on CPU

DGX agent

LFM2.5-Encoder is a family of multilingual bidirectional encoders built on the LFM2 architecture, available in two sizes: LFM2.5-Encoder-230M — a lightweight encoder for tight latency and memory budge

model-releasesr-localllama
28 Jul 2026
← Previous
1…354355356357358…1337
Next →