AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,860 results
Model Releases

FOREVER: Forgetting Curve-Inspired Memory Replay for Language Model Continual Learning

DGX agent

arXiv:2601.03938v2 Announce Type: replace-cross Abstract: Continual learning (CL) for large language models (LLMs) aims to enable sequential knowledge acquisition without catastrophic forgetting. Memo

model-releasesarxiv-cs-cl
21 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Applications

Grokking of Diffusion Models: Case Study on Modular Addition

DGX agent

arXiv:2604.17673v1 Announce Type: new Abstract: Despite their empirical success, how diffusion models generalize remains poorly understood from a mechanistic perspective. We demonstrate that diffusion

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

IncreFA: Breaking the Static Wall of Generative Model Attribution

DGX agent

arXiv:2604.17736v1 Announce Type: new Abstract: As AI generative models evolve at unprecedented speed, image attribution has become a moving target. New diffusion, adversarial and autoregressive gener

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

On the Predictive Power of Representation Dispersion in Language Models

DGX agent

arXiv:2506.24106v2 Announce Type: replace Abstract: We show that a language model's ability to predict text is tightly linked to the breadth of its embedding space: models that spread their contextual

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

One Adapts to Any: Meta Reward Modeling for Personalized LLM Alignment

DGX agent

arXiv:2601.18731v2 Announce Type: replace Abstract: Alignment of Large Language Models (LLMs) aims to align outputs with human preferences, and personalized alignment further adapts models to individu

safetyarxiv-cs-cl
21 Apr 2026
Research

SmoGVLM: A Small, Graph-enhanced Vision-Language Model

DGX agent

arXiv:2604.16517v1 Announce Type: cross Abstract: Large vision-language models (VLMs) achieve strong performance on multimodal tasks but often suffer from hallucination and poor grounding in knowledge

researcharxiv-cs-cl
21 Apr 2026
Model Releases

SpeakerSleuth: Can Large Audio-Language Models Judge Speaker Consistency across Multi-turn Dialogues?

DGX agent

arXiv:2601.04029v2 Announce Type: replace Abstract: Large Audio-Language Models (LALMs) as judges have emerged as a prominent approach for evaluating speech generation quality, yet their ability to as

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Stability-Weighted Decoding for Diffusion Language Models

DGX agent

arXiv:2604.17068v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) enable parallel text generation by iteratively denoising a fully masked sequence, unmasking a subset of masked t

researcharxiv-cs-cl
21 Apr 2026
Applications

TensorHub: Rethinking AI Model Hub with Tensor-Centric Compression

DGX agent

arXiv:2604.17104v1 Announce Type: cross Abstract: Modern AI models are growing rapidly in size and redundancy, leading to significant storage and distribution challenges in model hubs. We present Tens

applicationsarxiv-cs-lg
21 Apr 2026
Research

Advancing Intelligent Sequence Modeling: Evolution, Trade-offs, and Applications of State- Space Architectures from S4 to Mamba

DGX agent

arXiv:2503.18970v3 Announce Type: replace Abstract: Structured State Space Models (SSMs) have emerged as a transformative paradigm in sequence modeling, addressing critical limitations of Recurrent Ne

researcharxiv-cs-lg
20 Apr 2026
Model Releases

Do Vision-Language Models Truly Perform Vision Reasoning? A Rigorous Study of the Modality Gap

DGX agent

arXiv:2604.16256v1 Announce Type: cross Abstract: Reasoning in vision-language models (VLMs) has recently attracted significant attention due to its broad applicability across diverse downstream tasks

model-releasesarxiv-cs-cl
20 Apr 2026
Safety

Elucidating the SNR-t Bias of Diffusion Probabilistic Models

DGX agent

arXiv:2604.16044v1 Announce Type: new Abstract: Diffusion Probabilistic Models have demonstrated remarkable performance across a wide range of generative tasks. However, we have observed that these mo

safetyarxiv-cs-cv
20 Apr 2026
Research

LLaMo: Scaling Pretrained Language Models for Unified Motion Understanding and Generation with Continuous Autoregressive Tokens

DGX agent

arXiv:2602.12370v2 Announce Type: replace Abstract: Recent progress in large models has led to significant advances in unified multimodal generation and understanding. However, the development of mode

researcharxiv-cs-cv
20 Apr 2026
Model Releases

Moonshot AI releases Kimi-K2.6 model with 1T parameters, attention optimizations

DGX agent

Moonshot AI today released Kimi-K2.6, the latest addition to its popular Kimi series of open-source large language models. The Chinese artificial intelligence startup says that the algorithm outperfor

model-releasessiliconangle
20 Apr 2026
Model Releases

No Universal Courtesy: A Cross-Linguistic, Multi-Model Study of Politeness Effects on LLMs Using the PLUM Corpus

DGX agent

arXiv:2604.16275v1 Announce Type: new Abstract: This paper explores the response of Large Language Models (LLMs) to user prompts with different degrees of politeness and impoliteness. The Politeness T

model-releasesarxiv-cs-cl
20 Apr 2026
Research

Reward Modeling for Scientific Writing Evaluation

DGX agent

arXiv:2601.11374v2 Announce Type: replace Abstract: Scientific writing is an expert-domain task that demands deep domain knowledge, task-specific requirements and reasoning capabilities that leverage

researcharxiv-cs-cl
20 Apr 2026
Applications

Sketching the Readout of Large Language Models for Scalable Data Attribution and Valuation

DGX agent

arXiv:2604.16197v1 Announce Type: new Abstract: Data attribution and valuation are critical for understanding data-model synergy for Large Language Models (LLMs), yet existing gradient-based methods s

applicationsarxiv-cs-lg
20 Apr 2026
Agents

Some Mac Mini and Mac Studio models are unavailable or facing up to 12-week wait times in the US, with analysts citing strong demand from AI agent power users (Nicole Nguyen/Wall Street Journal)

DGX agent

Nicole Nguyen / Wall Street Journal: Some Mac Mini and Mac Studio models are unavailable or facing up to 12-week wait times in the US, with analysts citing strong demand from AI agent power users — Th

agentstechmeme
18 Apr 2026
Local Ai

AIPC: Agent-Based Automation for AI Model Deployment with Qualcomm AI Runtime

DGX agent

arXiv:2604.14661v1 Announce Type: cross Abstract: Edge AI model deployment is a multi-stage engineering process involving model conversion, operator compatibility handling, quantization calibration, r

local-aiarxiv-cs-lg
17 Apr 2026
Model Releases

Correcting Suppressed Log-Probabilities in Language Models with Post-Transformer Adapters

DGX agent

arXiv:2604.14174v1 Announce Type: new Abstract: Alignment-tuned language models frequently suppress factual log-probabilities on politically sensitive topics despite retaining the knowledge in their h

model-releasesarxiv-cs-cl
17 Apr 2026
Research

IMPACTX: improving model performance by appropriately constraining the training with teacher explanations

DGX agent

arXiv:2502.12222v2 Announce Type: replace Abstract: The eXplainable Artificial Intelligence (XAI) research predominantly concentrates to provide explainations about AI model decisions, especially Deep

researcharxiv-cs-lg
17 Apr 2026
Model Releases

Label-efficient underwater species classification with logistic regression on frozen foundation model embeddings

DGX agent

arXiv:2604.00313v2 Announce Type: replace Abstract: Automated species classification from underwater imagery is bottlenecked by the cost of expert annotation, and supervised models trained on one data

model-releasesarxiv-cs-cv
17 Apr 2026
Safety

Reasoning Dynamics and the Limits of Monitoring Modality Reliance in Vision-Language Models

DGX agent

arXiv:2604.14888v1 Announce Type: new Abstract: Recent advances in vision language models (VLMs) offer reasoning capabilities, yet how these unfold and integrate visual and textual information remains

safetyarxiv-cs-cl
17 Apr 2026
Agents

To go deeper on our new Life Sciences model series, research lead @joyjiao12 and product lead Yunyun Wang joined @AndrewMayne on the OpenAI …

DGX agent

To go deeper on our new Life Sciences model series, research lead @joyjiao12 and product lead Yunyun Wang joined @AndrewMayne on the OpenAI Podcast to discuss how we’re building models for biology, dr

agentsopenai--x
17 Apr 2026
Research

Towards Faster Language Model Inference Using Mixture-of-Experts Flow Matching

DGX agent

arXiv:2604.15009v1 Announce Type: cross Abstract: Flow matching retains the generation quality of diffusion models while enabling substantially faster inference, making it a compelling paradigm for ge

researcharxiv-cs-lg
17 Apr 2026
Model Releases

🇺🇸 xAI just dropped a real-time speech-to-text model built for voice apps: high limits, multi-language support, the works. Priced at just …

DGX agent

🇺🇸 xAI just dropped a real-time speech-to-text model built for voice apps: high limits, multi-language support, the works. Priced at just 0.10–0.20/hour, crushing most competitors. Quiet release, loud

model-releaseselon-musk--x
17 Apr 2026
Local Ai

Can you use Ollama models with the Codex app on Windows?

DGX agent

Yes, Ollama models can be used with the Codex app on Windows. Ollama supports all major operating systems, including Windows , and open models can be used with OpenAI's Codex CLI through Ollama — Code

local-air-ollama
16 Apr 2026
Model Releases

Cracking the Code of Juxtaposition: Can AI Models Understand the Humorous Contradictions

DGX agent

arXiv:2405.19088v3 Announce Type: replace Abstract: Recent advancements in large multimodal language models have demonstrated remarkable proficiency across a wide range of tasks. Yet, these models sti

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Diffusion Sequence Models for Generative In-Context Meta-Learning of Robot Dynamics

DGX agent

arXiv:2604.13366v1 Announce Type: new Abstract: Accurate modeling of robot dynamics is essential for model-based control, yet remains challenging under distributional shifts and real-time constraints.

researcharxiv-cs-lg
16 Apr 2026
Research

FAST: A Synergistic Framework of Attention and State-space Models for Spatiotemporal Traffic Prediction

DGX agent

arXiv:2604.13453v1 Announce Type: new Abstract: Traffic forecasting requires modeling complex temporal dynamics and long-range spatial dependencies over large sensor networks. Existing methods typical

researcharxiv-cs-lg
16 Apr 2026
Research

Indexing Multimodal Language Models for Large-scale Image Retrieval

DGX agent

arXiv:2604.13268v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated strong cross-modal reasoning capabilities, yet their potential for vision-only tasks remain

researcharxiv-cs-cl
16 Apr 2026
Research

Neural Chain-of-Thought Search: Searching the Optimal Reasoning Path to Enhance Large Language Models

DGX agent

arXiv:2601.11340v2 Announce Type: replace Abstract: Chain-of-Thought reasoning has significantly enhanced the problem-solving capabilities of Large Language Models. Unfortunately, current models gener

researcharxiv-cs-cl
16 Apr 2026
Model Releases

TLoRA+: A Low-Rank Parameter-Efficient Fine-Tuning Method for Large Language Models

DGX agent

arXiv:2604.13368v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) aims to adapt pre-trained models to specific tasks using relatively small and domain-specific datasets. Among P

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Benchmarking Foundation Models with Retrieval-Augmented Generation in Olympic-Level Physics Problem Solving

DGX agent

arXiv:2510.00919v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) with foundation models has achieved strong performance across diverse tasks, but their capacity for exper

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Climate Model Tuning with Online Synchronization-Based Parameter Estimation

DGX agent

arXiv:2510.06180v2 Announce Type: replace-cross Abstract: In climate science, the tuning of climate models is a computationally intensive problem due to the combination of the high-dimensionality of t

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Don't Show Pixels, Show Cues: Unlocking Visual Tool Reasoning in Language Models via Perception Programs

DGX agent

arXiv:2604.12896v1 Announce Type: new Abstract: Multimodal language models (MLLMs) are increasingly paired with vision tools (e.g., depth, flow, correspondence) to enhance visual reasoning. However, d

model-releasesarxiv-cs-cv
15 Apr 2026
Hardware

Fast AI Model Partition for Split Learning over Edge Networks

DGX agent

arXiv:2507.01041v4 Announce Type: replace-cross Abstract: Split learning (SL) is a distributed learning paradigm that can enable computation-intensive artificial intelligence (AI) applications by part

hardwarearxiv-cs-ai
15 Apr 2026
Research

PILOT: Planning via Internalized Latent Optimization Trajectories for Large Language Models

DGX agent

arXiv:2601.19917v2 Announce Type: replace Abstract: Strategic planning is critical for multi-step reasoning, yet compact Large Language Models (LLMs) often lack the capacity to formulate global strate

researcharxiv-cs-cl
15 Apr 2026
Research

Retrievals Can Be Detrimental: Unveiling the Backdoor Vulnerability of Retrieval-Augmented Diffusion Models

DGX agent

arXiv:2501.13340v4 Announce Type: replace Abstract: Diffusion models (DMs) have recently demonstrated remarkable generation capability. However, their training generally requires huge computational re

researcharxiv-cs-cv
15 Apr 2026
Safety

Task Alignment: A simple and effective proxy for model merging in computer vision

DGX agent

arXiv:2604.12935v1 Announce Type: new Abstract: Efficiently merging several models fine-tuned for different tasks, but stemming from the same pretrained base model, is of great practical interest. Des

safetyarxiv-cs-cv
15 Apr 2026
Model Releases

When Self-Reference Fails to Close: Matrix-Level Dynamics in Large Language Models

DGX agent

arXiv:2604.12128v1 Announce Type: new Abstract: We investigate how self-referential inputs alter the internal matrix dynamics of large language models. Measuring 106 scalar metrics across up to 7 anal

model-releasesarxiv-cs-cl
15 Apr 2026
Tutorials

A Mechanistic Analysis of Looped Reasoning Language Models

DGX agent

arXiv:2604.11791v1 Announce Type: cross Abstract: Reasoning has become a central capability in large language models. Recent research has shown that reasoning performance can be improved by looping an

tutorialsarxiv-cs-ai
14 Apr 2026
Local Ai

Abliterated (uncensored) models

DGX agent

This r/ollama discussion covers 'abliterated' models — LLMs that have had their built-in refusal mechanisms removed through a technique called abliteration, allowing them to respond to prompts without

local-air-ollama
14 Apr 2026
Tutorials

ActDistill: General Action-Guided Self-Derived Distillation for Efficient Vision-Language-Action Models

DGX agent

arXiv:2511.18082v3 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models have shown impressive flexibility and generalization, yet their deployment in robotic manipulation remain

tutorialsarxiv-cs-cv
14 Apr 2026
Local Ai

Ambiguity Detection and Elimination in Automated Executable Process Modeling

DGX agent

arXiv:2604.10884v1 Announce Type: cross Abstract: Automated generation of executable Business Process Model and Notation (BPMN) models from natural-language specifications is increasingly enabled by l

local-aiarxiv-cs-ai
14 Apr 2026
Tutorials

Can Small Training Runs Reliably Guide Data Curation? Rethinking Proxy-Model Practice

DGX agent

arXiv:2512.24503v2 Announce Type: replace-cross Abstract: Data teams at frontier AI companies routinely train small proxy models to make critical decisions about pretraining data recipes for full-scal

tutorialsarxiv-cs-ai
14 Apr 2026
Agents

Everyone I know is switching over to hermes agent, in large part because it actually works with smaller open source models.

DGX agent

Nous Research's Hermes agent framework has been gaining significant adoption due to its compatibility and effectiveness with smaller open-source language models, making it accessible beyond large prop

agentsnous-research--x
14 Apr 2026
Local Ai

One-click LM Studio → Ollama model linker

DGX agent

This r/ollama post discusses a tool for easily linking models between LM Studio and Ollama without duplicating disk storage. Both Ollama and LM Studio are popular local LLM tools, but they store their

local-air-ollama
14 Apr 2026
← Previous
1…4041424344…1248
Next →