AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlog
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,138 results
Model Releases

Tight Sample Complexity for Low-Rank Adaptation: Matching Bounds and Rank Selection

DGX agent

arXiv:2607.27680v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) has become the standard mechanism for fine-tuning large pretrained models, yet its statistical properties remain only parti

model-releasesarxiv-cs-cl
31 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Using an AMD V620 workstation card for ComfyUI - success

DGX agent

A few weeks ago I posted about if it was worth using a V620 for Comfyui, and was told it likely wouldn't work, at least in Windows 11. And if it did, it would be far too slow and unusable. I decided t

model-releasesr-stablediffusion
31 Jul 2026
Model Releases

@ArtificialAnlys ok @openai gets it https://x.com/OpenAI/status/2082878156483219672

DGX agent

@ArtificialAnlys ok @openai gets it https://x.com/OpenAI/status/2082878156483219672 We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are

model-releasesswyx--x
30 Jul 2026
Model Releases

b10188

DGX agent

metal: fix memory unwire if model is freed without any GPU operations (#26082) metal: fix memory leak if model is freed without any GPU operations metal: run dummy work only if residency sets are used

model-releasesllama-cpp-releases
30 Jul 2026
Model Releases

Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution

DGX agent

arXiv:2607.26596v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have demonstrated remarkable capabilities by integrating visual and textual understanding within a unified tran

model-releasesarxiv-cs-cv
30 Jul 2026
Industry

Deploying Kimi K3 on AWS

DGX agent

Kimi K3 is a 2.8‑trillion‑parameter Mixture of Experts model released by Moonshot AI on July 27, 2026, with its weights publicly available for self-hosting. Its architecture—using Kimi Delta Attention

industryaws-ml-blog
30 Jul 2026
Model Releases

Enhancing Generative Information Extraction with Two-step Validation: A Product Attribute Use Case

DGX agent

arXiv:2607.26780v1 Announce Type: new Abstract: The ability of large language models (LLMs) to process and generate text has introduced potential for applications in information extraction (IE). While

model-releasesarxiv-cs-cl
30 Jul 2026
Safety

HiFloat4 Format for End-To-End Reinforcement Learning Post-Training of Large Language Models

DGX agent

arXiv:2607.26515v1 Announce Type: new Abstract: We present, to our knowledge, the first end-to-end FP4 RL post-training, in which both the rollout and training policies, including their forward and ba

safetyarxiv-cs-lg
30 Jul 2026
Research

Linguistic Monoculture in LLM-Assisted Language Use

DGX agent

arXiv:2607.27134v1 Announce Type: cross Abstract: Writing and communication are increasingly mediated by large language models (LLMs) that are being used to draft, revise and polish text. Although suc

researcharxiv-cs-cl
30 Jul 2026
Model Releases

Position: Evaluation Scores Are Perishable Knowledge Claims

DGX agent

arXiv:2607.26191v1 Announce Type: cross Abstract: Evaluation methodologies for language models increasingly combine multiple signals, from automated metrics and LLM-as-judge ratings to human assessmen

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

RAGuard: A Layered Defense Framework for Retrieval-Augmented Generation Systems Against Data Poisoning

DGX agent

arXiv:2607.26339v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems ground large language models (LLMs) in external corpora, but this reliance exposes them to corpus poisoning

model-releasesarxiv-cs-lg
30 Jul 2026
Safety

Self-Adaptive Learning and Model Predictive Control for Tracking Unknown Dynamics with No Regret

DGX agent

arXiv:2607.26370v1 Announce Type: cross Abstract: We propose a self-adaptive online learning for control method for tracking unknown target dynamics. The target dynamics can exhibit switching behavior

safetyarxiv-cs-lg
30 Jul 2026
Model Releases

The Reliability of LLMs for Medical Diagnosis: An Examination of Consistency, Manipulation, and Contextual Awareness

DGX agent

arXiv:2503.10647v2 Announce Type: replace Abstract: This study evaluated the diagnostic reliability of two Large Language Models (LLMs), Google Gemini 2.0 Flash and OpenAI ChatGPT-4o, across three dim

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Thinking Machines just released Inkling-Small: 276B total, 12B active. A faster Inkling that matches or beats its 975B big brother in many b…

DGX agent

Thinking Machines just released Inkling-Small: 276B total, 12B active. A faster Inkling that matches or beats its 975B big brother in many benchmarks. To test its speed, we plugged it into HF's speech

model-releasessoumith-chintala--x
30 Jul 2026
Model Releases

b10174

DGX agent

model: add NextN/MTP speculative decoding support for GLM_DSA (GLM-5.2) (#25980) model: add NextN/MTP speculative decoding support for GLM_DSA (GLM-5.2) Adds GLM-5.2 NextN/MTP as a --spec-type draft-m

model-releasesllama-cpp-releases
29 Jul 2026
Model Releases

Beyond Static Costs: Learning-Dynamics Aware Loss Functions for Long-Tailed Classification

DGX agent

arXiv:2607.25830v1 Announce Type: new Abstract: Deep learning models in computer vision face significant challenges when trained on long-tailed datasets, where a few majority classes dominate while ma

model-releasesarxiv-cs-cv
29 Jul 2026
Safety

Contrastive Weak-to-strong Generalization

DGX agent

arXiv:2510.07884v3 Announce Type: replace-cross Abstract: Weak-to-strong generalization provides a promising paradigm for scaling large language models (LLMs) by training stronger models on samples fr

safetyarxiv-cs-ai
29 Jul 2026
Model Releases

HANDBOOK.md: A Benchmark for Long-Context Agentic Instruction Following

DGX agent

arXiv:2607.25398v1 Announce Type: new Abstract: Language-model agents are increasingly deployed under standing instructions: a system prompt, a policy file, or a skills document is placed in context,

model-releasesarxiv-cs-ai
29 Jul 2026
Research

KQFuzz: Knowledge-Guided Fuzzing for Quantum Libraries via Large Language Models

DGX agent

arXiv:2607.25647v1 Announce Type: cross Abstract: As quantum computing continually improves, ensuring the reliability and correctness of quantum libraries has become increasingly critical. To this end

researcharxiv-cs-ai
29 Jul 2026
Model Releases

LaP-Forensics: Latent-Pixel Consistency Guided Multimodal Reasoning for Deepfake Detection

DGX agent

arXiv:2607.25962v1 Announce Type: new Abstract: Recent generative models can produce images with few obvious visual artifacts, weakening detectors and explanations that rely only on surface appearance

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Localized Adaptation Reveals Distinct Learning Signatures in Transformers

DGX agent

arXiv:2607.25663v1 Announce Type: new Abstract: Transformer adaptation is typically distributed across model depth, even when the intended change is narrow. We investigate how adaptation site shapes w

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

M^2PO: Multi-Perspective Multi-Pair Preference Optimization for Machine Translation

DGX agent

arXiv:2510.13434v2 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with human preferences is pivotal for Machine Translation (MT), yet current approaches are often hindered by m

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

Multimodal Hybrid Retrieval-Augmented Generation for Scientific Document Understanding using Open-Source SLMs

DGX agent

arXiv:2607.24799v1 Announce Type: cross Abstract: Large Language Models tend to hallucinate when answering domain-specific ques tions from scientific documents without prior fine-tuning. Currently, me

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Parallel Decoding Distillation for Fast Image and Video Generation

DGX agent

arXiv:2607.26004v1 Announce Type: new Abstract: Generation in video diffusion or flow models is computationally expensive due to the slow and iterative sampling process. Current state-of-the-art (SOTA

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents

DGX agent

arXiv:2607.25485v1 Announce Type: new Abstract: Health AI is evolving from answering questions to agentic systems that converse with patients, reason about health records, and act on their behalf. Pri

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Physics-Grounded Fluid Video Generation with a Simulation Dataset and Dual-Stream Optical-Flow Supervision

DGX agent

arXiv:2607.25321v1 Announce Type: new Abstract: Video diffusion models generate visually compelling content but routinely violate elementary physics when the subject involves fluids: liquid columns br

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

The borderless Lakehouse: Bring AWS, Databricks and Snowflake data to your AI agents

DGX agent

Today’s data lakehouse is no longer mere data repository, but increasingly a system of action, actively executing tasks via always-on, autonomous AI agents. Rather than waiting for static reports, the

model-releasesgoogle-cloud-ai
29 Jul 2026
Local Ai

Athena-Brain Technical Report: An Efficient Robot Brain for General Intelligence and Embodied Interaction

DGX agent

arXiv:2607.18985v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated remarkable capabilities in language understanding, reasoning, and world knowledge. As embodied agents

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Benchmarking LLMs for Verilog Design Flows

DGX agent

arXiv:2607.22759v1 Announce Type: cross Abstract: Large language models (LLMs) show promise in code generation, but their capabilities to produce correct, synthesizable hardware description language (

model-releasesarxiv-cs-lg
28 Jul 2026
Safety

Beyond a Global Norm: Personalizing Toxicity Sensitivity in Language Models Without Retraining

DGX agent

arXiv:2607.23175v1 Announce Type: cross Abstract: Reducing toxicity is often framed as a global alignment problem, yet perceptions of harmful language are subjective and context-dependent. We present

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

Beyond ICA: Identifiability by Symmetry Breaking

DGX agent

arXiv:2607.23182v1 Announce Type: cross Abstract: We prove the identifiability of deep generative models (DGMs) with piecewise-affine (PWA) decoders and Gaussian mixture model (GMM) priors, in a purel

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

CausalGate: Causal Importance Distillation for Transformer Module Pruning

DGX agent

arXiv:2607.22720v1 Announce Type: cross Abstract: Existing adaptive inference methods for Large Language Models rely on observational heuristics, such as hidden-state similarity or activation magnitud

model-releasesarxiv-cs-cl
28 Jul 2026
Research

Differencing the Diffusion Trajectory toward Uncertain Components for Time Series Forecasting

DGX agent

arXiv:2607.22599v1 Announce Type: new Abstract: Diffusion models have become a widely used framework for probabilistic time series forecasting, modeling the distribution of future values given an obse

researcharxiv-cs-ai
28 Jul 2026
Model Releases

DuoAD: Leveraging [CLS] Dual Characteristics for Training-Free Few-Shot Anomaly Detection

DGX agent

arXiv:2607.23924v1 Announce Type: cross Abstract: Vision foundation models have enabled strong training-free anomaly detection (AD). However, most existing approaches rely primarily on independent loc

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

DynaCalKV: Key-Value Cache Compression via Head Grouping and Adaptive Rank Allocation

DGX agent

arXiv:2607.24331v1 Announce Type: new Abstract: As the inference phase of Large Language Models (LLMs) requires handling long context windows, the Key-Value (KV) cache initially appears to address thi

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

E-Bench: Benchmarking Multi-Step Tool-Use Agents in Real-World Product Scenarios

DGX agent

arXiv:2607.23722v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as agents that interact with stateful environments over multiple steps: gathering hidden informat

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Evidence Attribution in Visual Document Understanding without Coordinates or Region Labels

DGX agent

arXiv:2607.24651v1 Announce Type: cross Abstract: Reliable visual document understanding requires a model to attribute each answer to the evidence regions that support it. Recent benchmarks and system

researcharxiv-cs-cl
28 Jul 2026
Model Releases

Harmonized Interpretable ECG Waveform Features for Robust Cross-Dataset Clinical Prediction

DGX agent

arXiv:2607.23412v1 Announce Type: new Abstract: Electrocardiograms (ECGs) are widely used for cardiovascular risk prediction, yet models often fail to transfer across hospitals because of protocol, po

model-releasesarxiv-cs-lg
28 Jul 2026
Research

INSIGHT: Spatially resolved survival modelling from routine histology crosslinked with molecular profiling reveals prognostic epithelial-immune axes in stage II/III colorectal cancer

DGX agent

arXiv:2512.22262v2 Announce Type: replace-cross Abstract: Routine histology contains rich prognostic information in stage II/III colorectal cancer, much of which is embedded in complex spatial tissue

researcharxiv-cs-lg
28 Jul 2026
Model Releases

LFM2.5-Encoders: Fast at Long Context, Even on CPU

DGX agent

LFM2.5-Encoder is a family of multilingual bidirectional encoders built on the LFM2 architecture, available in two sizes: LFM2.5-Encoder-230M — a lightweight encoder for tight latency and memory budge

model-releasesr-localllama
28 Jul 2026
Tutorials

Like a Baby: Visually Situated Neural Language Acquisition

DGX agent

arXiv:1805.11546v3 Announce Type: replace-cross Abstract: We examine the benefits of visual context in training neural language models to perform next-word prediction. A multi-modal neural architectur

tutorialsarxiv-cs-ai
28 Jul 2026
Research

Not Forgotten: Implementation and Evaluation of a Personalized Episodic Memory for the Humanoid Robot Head Kim

DGX agent

arXiv:2607.24190v1 Announce Type: cross Abstract: Social robots that rely on large language models for conversation are unable to retain information across sessions. This absence of memory violates so

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Nova3D: Code-Native Generation of Programmable 3D Assets

DGX agent

arXiv:2607.22738v1 Announce Type: cross Abstract: Current 3D generative models mostly produce a final surface: a visually strong but largely opaque mesh. Interactive 3D worlds need more than a surface

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

ObsDriveBench: Benchmarking Multimodal Understanding under Adverse Weather with Observability Awareness

DGX agent

arXiv:2607.23537v1 Announce Type: new Abstract: Autonomous driving under adverse weather remains a critical challenge, yet existing vision-language benchmarks mainly evaluate under standard conditions

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

PANOPTICON: A PII-Based Assemblage of Naturalistic Output Tokens for Investigating Privacy Leakage Within LLM Context Window

DGX agent

arXiv:2607.22695v1 Announce Type: new Abstract: Large Language Models (LLMs) are capable of generalizing human language for the completion of never-before-seen tasks, leading to widespread deployment.

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Scale Weight Decay and Train Better

DGX agent

arXiv:2607.23777v1 Announce Type: cross Abstract: The discovery of scaling laws has motivated training neural networks on ever increasing quantities of data. This is typically done with a constant dec

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Seesaw: Accelerating Training by Balancing Learning Rate and Batch Size Scheduling

DGX agent

arXiv:2510.14717v2 Announce Type: replace-cross Abstract: Increasing the batch size during training -- a ''batch ramp'' -- is a promising strategy to accelerate large language model pretraining. While

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Spatial Reasoning in LLM Game Agents: Impact of Causal Context and Multi-Step Planning

DGX agent

arXiv:2607.22732v1 Announce Type: new Abstract: LLM-based game agents often perform poorly on more complex tasks. This work examines whether these failures are linked to limited spatial reasoning and

model-releasesarxiv-cs-ai
28 Jul 2026
← Previous
1…348349350351352…1316
Next →