AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,503 results
Safety

S-GRPO: Unified Post-Training for Large Vision-Language Models

DGX agent

arXiv:2604.16557v1 Announce Type: cross Abstract: Current post-training methodologies for adapting Large Vision-Language Models (LVLMs) generally fall into two paradigms: Supervised Fine-Tuning (SFT)

safetyarxiv-cs-cl
21 Apr 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Spatiotemporal Sycophancy: Negation-Based Gaslighting in Video Large Language Models

DGX agent

arXiv:2604.17873v1 Announce Type: new Abstract: Video Large Language Models (Vid-LLMs) have demonstrated remarkable performance in video understanding tasks, yet their robustness under conversational

model-releasesarxiv-cs-cv
21 Apr 2026
Research

TGLF-WINN: Data-Efficient Deep Learning Surrogate for Turbulent Transport Modeling in Fusion

DGX agent

arXiv:2509.07024v2 Announce Type: replace-cross Abstract: The Trapped Gyro-Landau Fluid (TGLF) model provides fast, accurate predictions of turbulent transport in tokamaks, but whole device simulation

researcharxiv-cs-lg
21 Apr 2026
Safety

Training Language Models to Use Prolog as a Tool

DGX agent

arXiv:2512.07407v2 Announce Type: replace Abstract: Language models frequently produce plausible yet incorrect reasoning traces that are difficult to verify. We investigate fine-tuning models to use P

safetyarxiv-cs-cl
21 Apr 2026
Safety

UniComp: A Unified Evaluation of Large Language Model Compression via Pruning, Quantization and Distillation

DGX agent

arXiv:2602.09130v3 Announce Type: replace Abstract: Model compression is increasingly essential for deploying large language models (LLMs), yet existing comparative studies largely focus on pruning an

safetyarxiv-cs-lg
21 Apr 2026
Research

Using Perspectival Words Is Harder Than Vocabulary Words for Humans and Even More So for Multimodal Language Models

DGX agent

arXiv:2506.00065v2 Announce Type: replace Abstract: Multimodal language models (MLMs) increasingly demonstrate human-like communication, yet their use of everyday perspectival words remains poorly und

researcharxiv-cs-cl
21 Apr 2026
Research

Vision Language Models are Biased

DGX agent

arXiv:2505.23941v4 Announce Type: replace-cross Abstract: Large language models (LLMs) memorize a vast amount of prior knowledge from the Internet that helps them on downstream tasks but also may noto

researcharxiv-cs-cv
21 Apr 2026
Safety

Where Do Self-Supervised Speech Models Become Unfair?

DGX agent

arXiv:2604.18249v1 Announce Type: new Abstract: Speech encoder models are known to model members of some speaker groups (SGs) better than others. However, there has been little work in establishing wh

safetyarxiv-cs-cl
21 Apr 2026
Research

LLMbench: A Comparative Close Reading Workbench for Large Language Models

DGX agent

arXiv:2604.15508v1 Announce Type: cross Abstract: LLMbench is a browser-based workbench for the comparative close reading of large language model (LLM) outputs. Where existing tools for LLM comparison

researcharxiv-cs-ai
20 Apr 2026
Research

Mechanisms of Prompt-Induced Hallucination in Vision-Language Models

DGX agent

arXiv:2601.05201v2 Announce Type: replace-cross Abstract: Large vision-language models (VLMs) are highly capable, yet often hallucinate by favoring textual prompts over visual evidence. We study this

researcharxiv-cs-ai
20 Apr 2026
Model Releases

MM-Telco: Benchmarks and Multimodal Large Language Models for Telecom Applications

DGX agent

arXiv:2511.13131v2 Announce Type: replace Abstract: Large Language Models (LLMs) have emerged as powerful tools for automating complex reasoning and decision-making tasks. In telecommunications, they

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

MTR-DuplexBench: Towards a Comprehensive Evaluation of Multi-Round Conversations for Full-Duplex Speech Language Models

DGX agent

arXiv:2511.10262v3 Announce Type: replace-cross Abstract: Full-Duplex Speech Language Models (FD-SLMs) enable real-time, overlapping conversational interactions, offering a more dynamic user experienc

model-releasesarxiv-cs-ai
20 Apr 2026
Agents

Security Threat Modeling for Emerging AI-Agent Protocols: A Comparative Analysis of MCP, A2A, Agora, and ANP

DGX agent

arXiv:2602.11327v2 Announce Type: replace-cross Abstract: The rapid development of the AI agent communication protocols, including the Model Context Protocol (MCP), Agent2Agent (A2A), Agora, and Agent

agentsarxiv-cs-ai
20 Apr 2026
Model Releases

TwoHamsters: Benchmarking Multi-Concept Compositional Unsafety in Text-to-Image Models

DGX agent

arXiv:2604.15967v1 Announce Type: cross Abstract: Despite the remarkable synthesis capabilities of text-to-image (T2I) models, safeguarding them against content violations remains a persistent challen

model-releasesarxiv-cs-cv
20 Apr 2026
Hardware

We’re launching Kimi K2.6 on Fireworks as a Day-0 launch partner! K2.5 was the base for standout models like @Cursor’s Composer 2 and was th…

DGX agent

We’re launching Kimi K2.6 on Fireworks as a Day-0 launch partner! K2.5 was the base for standout models like @Cursor’s Composer 2 and was the most popular model on our training platform. K2.6 on Firew

hardwarefireworks-ai--x
20 Apr 2026
Local Ai

@openclaw And of course @Ollama for the local model-serving engine. 🦙

DGX agent

Ollama is a local model-serving engine that enables users to run large language models on their own hardware without relying on cloud services. The post appears to highlight Ollama's integration with

local-aiollama--x
18 Apr 2026
Tools

Building a Fast Multilingual OCR Model with Synthetic Data

DGX agent

This article describes techniques for developing an efficient optical character recognition (OCR) model capable of processing multiple languages, leveraging synthetic data generation to reduce annotat

toolshugging-face
17 Apr 2026
Research

Contextuality from Single-State Ontological Models: An Information-Theoretic Obstruction

DGX agent

arXiv:2602.16716v3 Announce Type: replace Abstract: Contextuality is a central feature of quantum theory, traditionally understood as the impossibility of reproducing quantum measurement statistics us

researcharxiv-cs-ai
17 Apr 2026
Research

Edge-preserving noise for diffusion models

DGX agent

arXiv:2410.01540v4 Announce Type: replace Abstract: Classical diffusion models typically rely on isotropic Gaussian noise, treating all regions uniformly and overlooking structural information importa

researcharxiv-cs-cv
17 Apr 2026
Model Releases

Fact4ac at the Financial Misinformation Detection Challenge Task: Reference-Free Financial Misinformation Detection via Fine-Tuning and Few-Shot Prompting of Large Language Models

DGX agent

arXiv:2604.14640v1 Announce Type: new Abstract: The proliferation of financial misinformation poses a severe threat to market stability and investor trust, misleading market behavior and creating crit

model-releasesarxiv-cs-cl
17 Apr 2026
Research

GUI-Perturbed: Domain Randomization Reveals Systematic Brittleness in GUI Grounding Models

DGX agent

arXiv:2604.14262v1 Announce Type: new Abstract: GUI grounding models report over 85% accuracy on standard benchmarks, yet drop 27-56 percentage points when instructions require spatial reasoning rathe

researcharxiv-cs-lg
17 Apr 2026
Local Ai

How to Disable Thinking mode of Ollama Models Using Copilot CLI?

DGX agent

A guide on disabling thinking mode in Ollama models using the CLI by running models with the `--think=false` flag or using `/set nothink` followed by a prompt. The post likely discusses how to configu

local-air-ollama
17 Apr 2026
Agents

In-Context Autonomous Network Incident Response: An End-to-End Large Language Model Agent Approach

DGX agent

arXiv:2602.13156v2 Announce Type: replace-cross Abstract: Rapidly evolving cyberattacks demand incident response systems that can autonomously learn and adapt to changing threats. Prior work has exten

agentsarxiv-cs-ai
17 Apr 2026
Agents

Interpretable and Explainable Surrogate Modeling for Simulations: A State-of-the-Art Survey and Perspectives on Explainable AI for Decision-Making

DGX agent

arXiv:2604.14240v1 Announce Type: cross Abstract: The simulation of complex systems increasingly relies on sophisticated but fundamentally opaque computational black-box simulators. Surrogate models p

agentsarxiv-cs-lg
17 Apr 2026
Safety

Language on Demand, Knowledge at Core: Composing LLMs with Encoder-Decoder Translation Models for Extensible Multilinguality

DGX agent

arXiv:2603.17512v4 Announce Type: replace Abstract: Large language models (LLMs) exhibit strong general intelligence, yet their multilingual performance remains highly imbalanced. Although LLMs encode

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

LLMOrbit: A Circular Taxonomy of Large Language Models -From Scaling Walls to Agentic AI Systems

DGX agent

arXiv:2601.14053v2 Announce Type: replace-cross Abstract: The field of artificial intelligence has undergone a revolution from foundational Transformer architectures to reasoning-capable systems appro

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

MetaDent: Labeling Clinical Images for Vision-Language Models in Dentistry

DGX agent

arXiv:2604.14866v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated significant potential in medical image analysis, yet their application in intraoral photography remains

model-releasesarxiv-cs-cv
17 Apr 2026
Safety

SPAGBias: Uncovering and Tracing Structured Spatial Gender Bias in Large Language Models

DGX agent

arXiv:2604.14672v1 Announce Type: new Abstract: Large language models (LLMs) are being increasingly used in urban planning, but since gendered space theory highlights how gender hierarchies are embedd

safetyarxiv-cs-cl
17 Apr 2026
Safety

Switch-KD: Visual-Switch Knowledge Distillation for Vision-Language Models

DGX agent

arXiv:2604.14629v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have shown remarkable capabilities in joint vision-language understanding, but their large scale poses significant challen

safetyarxiv-cs-cv
17 Apr 2026
Safety

Towards Fine-grained Temporal Perception: Post-Training Large Audio-Language Models with Audio-Side Time Prompt

DGX agent

arXiv:2604.13715v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) enable general audio understanding and demonstrate remarkable performance across various audio tasks. However, the

safetyarxiv-cs-ai
17 Apr 2026
Model Releases

Weight Patching: Toward Source-Level Mechanistic Localization in LLMs

DGX agent

arXiv:2604.13694v1 Announce Type: new Abstract: Mechanistic interpretability seeks to localize model behavior to the internal components that causally realize it. Prior work has advanced activation-sp

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

What folk don’t get is that the play of Anthropic et al is to get enterprises to use their models via their own wrapper hooked into systems …

DGX agent

What folk don’t get is that the play of Anthropic et al is to get enterprises to use their models via their own wrapper hooked into systems of record This is why they are moving from per seat pricing

model-releasesemad-mostaque--x
17 Apr 2026
Tutorials

World-Value-Action Model: Implicit Planning for Vision-Language-Action Systems

DGX agent

arXiv:2604.14732v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for building embodied agents that ground perception and language into action.

tutorialsarxiv-cs-lg
17 Apr 2026
Model Releases

Auto-FP: An Experimental Study of Automated Feature Preprocessing for Tabular Data

DGX agent

arXiv:2310.02540v2 Announce Type: replace Abstract: Classical machine learning models, such as linear models and tree-based models, are widely used in industry. These models are sensitive to data dist

model-releasesarxiv-cs-lg
16 Apr 2026
Safety

Beyond State Consistency: Behavior Consistency in Text-Based World Models

DGX agent

arXiv:2604.13824v1 Announce Type: new Abstract: World models have been emerging as critical components for assessing the consequences of actions generated by interactive agents in online planning and

safetyarxiv-cs-lg
16 Apr 2026
Model Releases

Can Large Language Models Reliably Extract Physiology Index Values from Coronary Angiography Reports?

DGX agent

arXiv:2604.13077v1 Announce Type: new Abstract: Coronary angiography (CAG) reports contain clinically relevant physiological measurements, yet this information is typically in the form of unstructured

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Evaluating Supervised Machine Learning Models: Principles, Pitfalls, and Metric Selection

DGX agent

arXiv:2604.13882v1 Announce Type: new Abstract: The evaluation of supervised machine learning models is a critical stage in the development of reliable predictive systems. Despite the widespread avail

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

F-Actor: Controllable Conversational Behaviour in Full-Duplex Models

DGX agent

arXiv:2601.11329v3 Announce Type: replace Abstract: Spoken conversational systems require more than accurate speech generation to have human-like conversations: to feel natural and engaging, they must

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Free Lunch for Unified Multimodal Models: Enhancing Generation via Reflective Rectification with Inherent Understanding

DGX agent

arXiv:2604.13540v1 Announce Type: new Abstract: Unified Multimodal Models (UMMs) aim to integrate visual understanding and generation within a single structure. However, these models exhibit a notable

researcharxiv-cs-cv
16 Apr 2026
Model Releases

Models are getting smaller, smarter and Apache licensed. Love to see Gemma and Qwen doing it.

DGX agent

Models are getting smaller, smarter and Apache licensed. Love to see Gemma and Qwen doing it. Qwen 3.6 is here, and open-source! Run it locally with improved agentic coding capabilities. Try it with C

model-releasesollama--x
16 Apr 2026
Agents

POINTS-Seeker: Towards Training a Multimodal Agentic Search Model from Scratch

DGX agent

arXiv:2604.14029v1 Announce Type: new Abstract: While Large Multimodal Models (LMMs) demonstrate impressive visual perception, they remain epistemically constrained by their static parametric knowledg

agentsarxiv-cs-cv
16 Apr 2026
Model Releases

2/5 Turns out the model wasn't remembering the solution, but it was identifying 'gold-like' aesthetics like minimality & clarity. Total form…

DGX agent

AI21 Labs shared findings indicating that their model does not simply memorize solutions but instead identifies and recognizes aesthetic qualities associated with high-quality outputs, such as minimal

model-releasesai21-labs--x
15 Apr 2026
Model Releases

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction

DGX agent

arXiv:2603.07083v2 Announce Type: replace Abstract: Model-based reinforcement learning (MBRL) agents operating in high-dimensional observation spaces, such as Dreamer, rely on learning abstract repres

model-releasesarxiv-cs-lg
15 Apr 2026
Research

E2LLM: Encoder Elongated Large Language Models for Long-Context Understanding and Reasoning

DGX agent

arXiv:2409.06679v3 Announce Type: replace Abstract: Processing long contexts is increasingly important for Large Language Models (LLMs) in tasks like multi-turn dialogues, code generation, and documen

researcharxiv-cs-cl
15 Apr 2026
Model Releases

Guide to prompting Gemini 3.1 Flash TTS (text-to-speech)

DGX agent

Today, Gemini 3.1 Flash TTS, our latest text-to-speech model, is available on Google AI Studio and Vertex AI. It delivers precise controllability and expressivity, empowering developers and enterprise

model-releasesgoogle-cloud-ai
15 Apr 2026
Tutorials

KoCo: Conditioning Language Model Pre-training on Knowledge Coordinates

DGX agent

arXiv:2604.12397v1 Announce Type: new Abstract: Standard Large Language Model (LLM) pre-training typically treats corpora as flattened token sequences, often overlooking the real-world context that hu

tutorialsarxiv-cs-cl
15 Apr 2026
Model Releases

KumoRFM-2: Scaling Foundation Models for Relational Learning

DGX agent

arXiv:2604.12596v1 Announce Type: cross Abstract: We introduce KumoRFM-2, the next iteration of a pre-trained foundation model for relational data. KumoRFM-2 supports in-context learning as well as fi

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Large Language Models are Powerful Electronic Health Record Encoders

DGX agent

arXiv:2502.17403v5 Announce Type: replace-cross Abstract: Electronic Health Records (EHRs) offer considerable potential for clinical prediction, but their complexity and heterogeneity challenge tradit

model-releasesarxiv-cs-ai
15 Apr 2026
← Previous
1…9293949596…1261
Next →