AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
92,405Total entries
1Added by human
92,404Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,927 results
Research

DiT as Real-Time Rerenderer: Streaming Video Stylization with Autoregressive Diffusion Transformer

DGX agent

arXiv:2604.13509v1 Announce Type: new Abstract: Recent advances in video generation models has significantly accelerated video generation and related downstream tasks. Among these, video stylization h

researcharxiv-cs-cv
16 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Enhancing Mixture-of-Experts Specialization via Cluster-Aware Upcycling

DGX agent

arXiv:2604.13508v1 Announce Type: new Abstract: Sparse Upcycling provides an efficient way to initialize a Mixture-of-Experts (MoE) model from pretrained dense weights instead of training from scratch

researcharxiv-cs-cv
16 Apr 2026
Model Releases

Evaluating LLM-Based Translation of a Low-Resource Technical Language: The Medical and Philosophical Greek of Galen

DGX agent

arXiv:2602.24119v2 Announce Type: replace Abstract: Purpose: This study evaluates the quality of commercial large language model (LLM) machine translation (MT) for Ancient Greek technical prose and be

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

EVE: A Domain-Specific LLM Framework for Earth Intelligence

DGX agent

arXiv:2604.13071v1 Announce Type: new Abstract: We introduce Earth Virtual Expert (EVE), the first open-source, end-to-end initiative for developing and deploying domain-specialized LLMs for Earth Int

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Gemini can now pull from Google Photos to generate personalized images

DGX agent

Google's Personal Intelligence feature, which lets Gemini pull data from apps like Google Photos to offer responses tailored to you, can now use that data and its Nano Banana 2 image model to create i

model-releasesthe-verge-ai
16 Apr 2026
Applications

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System

DGX agent

arXiv:2604.14125v1 Announce Type: new Abstract: While end-to-end Vision-Language-Action (VLA) models offer a promising paradigm for robotic manipulation, fine-tuning them on narrow control data often

applicationsarxiv-cs-cv
16 Apr 2026
Safety

I have found that asking for a sestina regularly triggers Opus 4.7's safety guardrails. The forbidden poetic form!

DGX agent

Ethan Mollick reported that requesting Claude Opus 4.7 to write sestinas—a complex poetic form with strict structural requirements—frequently triggers the model's safety guardrails, suggesting the AI

safetyethan-mollick--x
16 Apr 2026
Model Releases

MAny: Merge Anything for Multimodal Continual Instruction Tuning

DGX agent

arXiv:2604.14016v1 Announce Type: new Abstract: Multimodal Continual Instruction Tuning (MCIT) is essential for sequential task adaptation of Multimodal Large Language Models (MLLMs) but is severely r

model-releasesarxiv-cs-lg
16 Apr 2026
Safety

Med-CAM: Minimal Evidence for Explaining Medical Decision Making

DGX agent

arXiv:2604.13695v1 Announce Type: new Abstract: Reliable and interpretable decision-making is essential in medical imaging, where diagnostic outcomes directly influence patient care. Despite advances

safetyarxiv-cs-cv
16 Apr 2026
Model Releases

MERRIN: A Benchmark for Multimodal Evidence Retrieval and Reasoning in Noisy Web Environments

DGX agent

arXiv:2604.13418v1 Announce Type: new Abstract: Motivated by the underspecified, multi-hop nature of search queries and the multimodal, heterogeneous, and often conflicting nature of real-world web re

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

More on my blog, including results from the previously secret 'flamingo on a unicycle' test https://simonwillison.net/2026/Apr/16/qwen-beats…

DGX agent

Simon Willison discusses results from a 'flamingo on a unicycle' test on his blog, likely comparing AI model performance including Qwen. The post appears to reference previously undisclosed or unconve

model-releasessimon-willison--x
16 Apr 2026
Local Ai

One Token per Highly Selective Frame: Towards Extreme Compression for Long Video Understanding

DGX agent

arXiv:2604.14149v1 Announce Type: new Abstract: Long video understanding is inherently challenging for vision-language models (VLMs) because of the extensive number of frames. With each video frame ty

local-aiarxiv-cs-cv
16 Apr 2026
Research

OneHOI: Unifying Human-Object Interaction Generation and Editing

DGX agent

arXiv:2604.14062v1 Announce Type: new Abstract: Human-Object Interaction (HOI) modelling captures how humans act upon and relate to objects, typically expressed as triplets. Existing approaches split

researcharxiv-cs-cv
16 Apr 2026
Model Releases

OPTED: Open Preprocessed Trachoma Eye Dataset Using Zero-Shot SAM 3 Segmentation

DGX agent

arXiv:2603.06885v2 Announce Type: replace Abstract: Trachoma remains the leading infectious cause of blindness worldwide, with Sub-Saharan Africa bearing over 85% of the global burden and Ethiopia alo

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Opus 4.7 is now supported in Hermes Agent 🚀🚀

DGX agent

Opus 4.7 is now supported in Hermes Agent 🚀🚀 Introducing Claude Opus 4.7, our most capable Opus model yet. It handles long-running tasks with more rigor, follows instructions more precisely, and verif

model-releasesnous-research--x
16 Apr 2026
Model Releases

Out of Context: Reliability in Multimodal Anomaly Detection Requires Contextual Inference

DGX agent

arXiv:2604.13252v1 Announce Type: new Abstract: Anomaly detection aims to identify observations that deviate from expected behavior. Because anomalous events are inherently sparse, most frameworks are

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Red Skills or Blue Skills? A Dive Into Skills Published on ClawHub

DGX agent

arXiv:2604.13064v1 Announce Type: new Abstract: Skill ecosystems have emerged as an increasingly important layer in Large Language Model (LLM) agent systems, enabling reusable task packaging, public d

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Replit Agent 4 is even smarter now with Claude Opus 4.7! 50% off for a limited time. Go try it now ↓

DGX agent

Replit Agent 4 has been upgraded to use Claude Opus 4.7, an advanced AI model, enhancing its code generation and problem-solving capabilities. The company is offering a 50% discount for a limited time

model-releasesreplit--x
16 Apr 2026
Model Releases

Response.

DGX agent

Response. Hey Ethan! Sean here, PM on http://Claude.ai - thanks for the feedback. This isn't a router, this is the model being trained to decide when to think based on the context -- we've been runnin

model-releasesethan-mollick--x
16 Apr 2026
Local Ai

Rhetorical Questions in LLM Representations: A Linear Probing Study

DGX agent

arXiv:2604.14128v1 Announce Type: new Abstract: Rhetorical questions are asked not to seek information but to persuade or signal stance. How large language models internally represent them remains unc

local-aiarxiv-cs-cl
16 Apr 2026
Research

Robust Ultra Low-Bit Post-Training Quantization via Stable Diagonal Curvature Estimate

DGX agent

arXiv:2604.13806v1 Announce Type: new Abstract: Large Language Models (LLMs) are widely used across many domains, but their scale makes deployment challenging. Post-Training Quantization (PTQ) reduces

researcharxiv-cs-lg
16 Apr 2026
Model Releases

RPS: Information Elicitation with Reinforcement Prompt Selection

DGX agent

arXiv:2604.13817v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable capabilities in dialogue generation and reasoning, yet their effectiveness in eliciting user-known bu

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention

DGX agent

arXiv:2604.13847v1 Announce Type: new Abstract: While sparse attention mitigates the computational bottleneck of long-context LLM training, its distributed training process exhibits extreme heterogene

model-releasesarxiv-cs-lg
16 Apr 2026
Research

Stability Principle Underlying Passive Dynamic Walking of Rimless Wheel

DGX agent

arXiv:2604.13530v1 Announce Type: new Abstract: Rimless wheels are known as the simplest model for passive dynamic walking. It is known that the passive gait generated only by gravity effect always be

researcharxiv-cs-ro
16 Apr 2026
Model Releases

Syn-TurnTurk: A Synthetic Dataset for Turn-Taking Prediction in Turkish Dialogues

DGX agent

arXiv:2604.13620v1 Announce Type: new Abstract: Managing natural dialogue timing is a significant challenge for voice-based chatbots. Most current systems usually rely on simple silence detection, whi

model-releasesarxiv-cs-cl
16 Apr 2026
Syntheses

Synthesis: Arxiv-Cs-Cl

DGX agent

Auto-generated synthesis of 505 entries about arxiv-cs-cl

synthesisarxiv-cs-clauto-generated
16 Apr 2026
Model Releases

Towards Generalizable Robotic Manipulation in Dynamic Environments

DGX agent

arXiv:2603.15620v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models excel in static manipulation but struggle in dynamic environments with moving targets. This performance gap prim

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Treating enterprise AI as an operating layer

DGX agent

There’s a fault line running through enterprise AI, and it’s not the one getting the most attention. The public conversation still tracks foundation models and benchmarks—GPT versus Gemini, reasoning

model-releasesmit-tech-review
16 Apr 2026
Model Releases

VGGT-Segmentor: Geometry-Enhanced Cross-View Segmentation

DGX agent

arXiv:2604.13596v1 Announce Type: new Abstract: Instance-level object segmentation across disparate egocentric and exocentric views is a fundamental challenge in visual understanding, critical for app

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

We fixed a bug where rate limits on Claude subscriptions weren't properly adjusted for long context requests in Opus 4.7. We've reset 5-hour…

DGX agent

Anthropic fixed a bug in Claude Opus 4.7 where rate limits for paid subscriptions weren't correctly adjusted for requests using the model's extended context window capabilities. The fix involved reset

model-releasesboris-cherny--x
16 Apr 2026
Local Ai

A Gustav Klimt–style lora for flux

DGX agent

This r/StableDiffusion post showcases a community-created LoRA model built for the Flux image generation architecture, designed to replicate the distinctive artistic style of Gustav Klimt (1862–1918),

local-air-stablediffusion
15 Apr 2026
Local Ai

A Hybrid Architecture for Benign-Malignant Classification of Mammography ROIs

DGX agent

arXiv:2604.12437v1 Announce Type: new Abstract: Accurate characterization of suspicious breast lesions in mammography is important for early diagnosis and treatment planning. While Convolutional Neura

local-aiarxiv-cs-cv
15 Apr 2026
Model Releases

Beyond Relevance: On the Relationship Between Retrieval and RAG Information Coverage

DGX agent

arXiv:2603.08819v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) systems combine document retrieval with a generative model to address complex information seeking tasks l

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Built GPT-2, Llama 3, and DeepSeek from scratch in PyTorch - open source code + book [p]

DGX agent

A Reddit post on r/MachineLearning sharing an open-source project and accompanying book by Sebastian Raschka that walks through implementing GPT-2, Llama 3, and DeepSeek from scratch using PyTorch, wi

model-releasesr-machinelearning
15 Apr 2026
Research

Calibrated Confidence Estimation for Tabular Question Answering

DGX agent

arXiv:2604.12491v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed for tabular question answering, yet calibration on structured data is largely unstudied. This pap

researcharxiv-cs-cl
15 Apr 2026
Safety

Cloud CISO Perspectives: How CISOs can pursue technical and cultural resilience (Q&A)

DGX agent

Welcome to the first Cloud CISO Perspectives for April 2026. Today, Thiébaut Meyer and Lia Wertheimer from Google Cloud’s Office of the CISO share Thiébaut’s conversation with Matt Rowe, chief securit

safetygoogle-cloud-ai
15 Apr 2026
Research

Decidable By Construction: Design-Time Verification for Trustworthy AI

DGX agent

arXiv:2603.25414v2 Announce Type: replace-cross Abstract: A prevailing assumption in machine learning is that model correctness must be enforced after the fact. We observe that the properties determin

researcharxiv-cs-ai
15 Apr 2026
Research

DeepTest Tool Competition 2026: Benchmarking an LLM-Based Automotive Assistant

DGX agent

arXiv:2604.12615v1 Announce Type: new Abstract: This report summarizes the results of the first edition of the Large Language Model (LLM) Testing competition, held as part of the DeepTest workshop at

researcharxiv-cs-ai
15 Apr 2026
Research

Distorted or Fabricated? A Survey on Hallucination in Video LLMs

DGX agent

arXiv:2604.12944v1 Announce Type: cross Abstract: Despite significant progress in video-language modeling, hallucinations remain a persistent challenge in Video Large Language Models (Vid-LLMs), refer

researcharxiv-cs-ai
15 Apr 2026
Model Releases

Do VLMs Truly 'Read' Candlesticks? A Multi-Scale Benchmark for Visual Stock Price Forecasting

DGX agent

arXiv:2604.12659v1 Announce Type: cross Abstract: Vision-language models(VLMs) are increasingly applied to visual stock price forecasting, yet existing benchmarks inadequately evaluate their understan

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

DPC-VQA: Decoupling Quality Perception and Residual Calibration for Video Quality Assessment

DGX agent

arXiv:2604.12813v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs) have shown promising performance on video quality assessment (VQA) tasks. However, adapting them to new

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

From Plan to Action: How Well Do Agents Follow the Plan?

DGX agent

arXiv:2604.12147v1 Announce Type: cross Abstract: Agents aspire to eliminate the need for task-specific prompt crafting through autonomous reason-act-observe loops. Still, they are commonly instructed

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Generative Anonymization in Event Streams

DGX agent

arXiv:2604.12803v1 Announce Type: new Abstract: Neuromorphic vision sensors offer low latency and high dynamic range, but their deployment in public spaces raises severe data protection concerns. Rece

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

GF-Score: Certified Class-Conditional Robustness Evaluation with Fairness Guarantees

DGX agent

arXiv:2604.12757v1 Announce Type: cross Abstract: Adversarial robustness is essential for deploying neural networks in safety-critical applications, yet standard evaluation methods either require expe

model-releasesarxiv-cs-ai
15 Apr 2026
Local Ai

I 'made' a patch for ernie-image fp16 support in comfyui (for 20 series cards)

DGX agent

A Reddit user shared a community-made patch to enable proper FP16 support for the Ernie Image model in ComfyUI, targeting NVIDIA 20-series (Turing) GPUs. 20-series cards do not support bfloat16 , and

local-air-stablediffusion
15 Apr 2026
Model Releases

Identity as Attractor: Geometric Evidence for Persistent Agent Architecture in LLM Activation Space

DGX agent

arXiv:2604.12016v1 Announce Type: new Abstract: Large language models map semantically related prompts to similar internal representations -- a phenomenon interpretable as attractor-like dynamics. We

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

INDOTABVQA: A Benchmark for Cross-Lingual Table Understanding in Bahasa Indonesia Documents

DGX agent

arXiv:2604.11970v1 Announce Type: cross Abstract: We introduce INDOTABVQA, a benchmark for evaluating cross-lingual Table Visual Question Answering (VQA) on real-world document images in Bahasa Indone

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT

DGX agent

arXiv:2512.14732v2 Announce Type: replace-cross Abstract: Incidental findings in CT scans, though often benign, can have significant clinical implications and should be reported following established

model-releasesarxiv-cs-ai
15 Apr 2026
← Previous
1…568569570571572…1395
Next →