AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,537 results
Model Releases

Opus 4.7 is a model I’ve loved working with in Claude Code. It’s more agentic and instruction following but also incredibly smart and creati…

DGX agent

Opus 4.7 is a model I’ve loved working with in Claude Code. It’s more agentic and instruction following but also incredibly smart and creative. I think it takes a slight adjustment to get used to, but

model-releasesthariq--x
16 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Reward Design for Physical Reasoning in Vision-Language Models

DGX agent

arXiv:2604.13993v1 Announce Type: cross Abstract: Physical reasoning over visual inputs demands tight integration of visual perception, domain knowledge, and multi-step symbolic inference. Yet even st

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

UniBlendNet: Unified Global, Multi-Scale, and Region-Adaptive Modeling for Ambient Lighting Normalization

DGX agent

arXiv:2604.13383v1 Announce Type: new Abstract: Ambient Lighting Normalization (ALN) aims to restore images degraded by complex, spatially varying illumination conditions. Existing methods, such as IF

model-releasesarxiv-cs-cv
16 Apr 2026
Research

VLMs Need Words: Vision Language Models Ignore Visual Detail In Favor of Semantic Anchors

DGX agent

arXiv:2604.02486v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have achieved impressive performance across a wide range of multimodal tasks. However, they often fail on tasks

researcharxiv-cs-cl
16 Apr 2026
Model Releases

When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning?

DGX agent

arXiv:2503.23137v2 Announce Type: replace-cross Abstract: Understanding humor-particularly when it involves complex, contradictory narratives that require comparative reasoning-remains a significant c

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

A General Model for Deepfake Speech Detection: Diverse Bonafide Resources or Diverse AI-Based Generators

DGX agent

arXiv:2603.27557v2 Announce Type: replace-cross Abstract: In this paper, we analyze two main factors of Bonafide Resource (BR) or AI-based Generator (AG) which affect the performance and the generalit

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

ByteDance launches its Seedance 2.0 video model to enterprise clients in 100+ countries, excluding the US amid legal disputes, after a February launch in China (Juro Osawa/The Information)

DGX agent

Juro Osawa / The Information: ByteDance launches its Seedance 2.0 video model to enterprise clients in 100+ countries, excluding the US amid legal disputes, after a February launch in China — ByteDanc

model-releasestechmeme
15 Apr 2026
Research

Conflated Inverse Modeling to Generate Diverse and Temperature-Change Inducing Urban Vegetation Patterns

DGX agent

arXiv:2604.13028v1 Announce Type: new Abstract: Urban areas are increasingly vulnerable to thermal extremes driven by rapid urbanization and climate change. Traditionally, thermal extremes have been m

researcharxiv-cs-cv
15 Apr 2026
Research

Dynamic Modeling and Robust Gait Optimization of a Compliant Worm Robot

DGX agent

arXiv:2604.12031v1 Announce Type: new Abstract: Worm-inspired robots provide an effective locomotion strategy for constrained environments by combining cyclic body deformation with alternating anchori

researcharxiv-cs-ro
15 Apr 2026
Model Releases

GroupKAN: Efficient Kolmogorov-Arnold Networks via Grouped Spline Modeling

DGX agent

arXiv:2511.05477v2 Announce Type: replace Abstract: Medical image segmentation demands models that achieve high accuracy while maintaining computational efficiency and clinical interpretability. While

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

HazardArena: Evaluating Semantic Safety in Vision-Language-Action Models

DGX agent

arXiv:2604.12447v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models inherit rich world knowledge from vision-language backbones and acquire executable skills via action demonstrations.

model-releasesarxiv-cs-ro
15 Apr 2026
Applications

LLM-HYPER: Generative CTR Modeling for Cold-Start Ad Personalization via LLM-Based Hypernetworks

DGX agent

arXiv:2604.12096v1 Announce Type: new Abstract: On online advertising platforms, newly introduced promotional ads face the cold-start problem, as they lack sufficient user feedback for model training.

applicationsarxiv-cs-ai
15 Apr 2026
Safety

MODIX: A Training-Free Multimodal Information-Driven Positional Index Scaling for Vision-Language Models

DGX agent

arXiv:2604.12537v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved remarkable progress in multimodal understanding, yet their positional encoding mechanisms remain suboptima

safetyarxiv-cs-ai
15 Apr 2026
Model Releases

Narrative over Numbers: The Identifiable Victim Effect and its Amplification Under Alignment and Reasoning in Large Language Models

DGX agent

arXiv:2604.12076v1 Announce Type: cross Abstract: The Identifiable Victim Effect (IVE) - the tendency to allocate greater resources to a specific, narratively described victim than to a statistically

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: Professional Image Quality Assessment (Track 1)

DGX agent

arXiv:2604.12512v1 Announce Type: cross Abstract: In this paper, we present an overview of the NTIRE 2026 challenge on the 3rd Restore Any Image Model in the Wild, specifically focusing on Track 1: Pr

model-releasesarxiv-cs-ai
15 Apr 2026
Safety

Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models

DGX agent

arXiv:2604.12582v1 Announce Type: new Abstract: Recent Video Large Language Models (Video-LLMs) have demonstrated strong capability in video understanding, yet they still suffer from hallucinations. E

safetyarxiv-cs-cv
15 Apr 2026
Local Ai

Running a 31B model locally made me realize how insane LLM infra actually is

DGX agent

A Reddit post from r/ollama in which a user shares their experience running a 31B parameter model locally using Ollama, reflecting on the surprisingly demanding hardware and infrastructure requirement

local-air-ollama
15 Apr 2026
Model Releases

The example prompt for Google's new Gemini Flash TTS text-to-speed model is a lot https://simonwillison.net/2026/Apr/15/gemini-31-flash-tts/

DGX agent

Google's Gemini 3.1 Flash TTS (text-to-speech) model includes a notably elaborate or extensive example prompt, which Simon Willison highlighted as noteworthy. The post likely comments on the complexit

model-releasessimon-willison--x
15 Apr 2026
Model Releases

Today we launched Gemini 3.1 Flash TTS, our most expressive and controllable text-to-speech model yet. This launch [excitement] includes aud…

DGX agent

Today we launched Gemini 3.1 Flash TTS, our most expressive and controllable text-to-speech model yet. This launch [excitement] includes audio tags! 🗣🏷 Audio tags [explanatory] are a seamless way to g

model-releasesgoogle-ai--x
15 Apr 2026
Model Releases

Understanding or Memorizing? A Case Study of German Definite Articles in Language Models

DGX agent

arXiv:2601.09313v2 Announce Type: replace-cross Abstract: Language models perform well on grammatical agreement, but it is unclear whether this reflects rule-based generalization or memorization. We s

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Variation in Verification: Understanding Verification Dynamics in Large Language Models

DGX agent

arXiv:2509.17995v2 Announce Type: replace-cross Abstract: Recent advances have shown that scaling test-time computation enables large language models (LLMs) to solve increasingly complex problems acro

model-releasesarxiv-cs-ai
15 Apr 2026
Research

Visual Diffusion Models are Geometric Solvers

DGX agent

arXiv:2510.21697v2 Announce Type: replace Abstract: In this paper we show that visual diffusion models can serve as effective geometric solvers: they can directly reason about geometric problems by wo

researcharxiv-cs-cv
15 Apr 2026
Tools

What if you could get 1.3B Transformer quality from a 770M model? That's not a compression result. It's a different architecture. Parcae, fr…

DGX agent

What if you could get 1.3B Transformer quality from a 770M model? That's not a compression result. It's a different architecture. Parcae, from @realDanFu (Together AI's VP of Kernels) and his lab at U

toolstogether-ai--x
15 Apr 2026
Local Ai

Which cloud model do you use for coding? Which one got better reasoning?

DGX agent

This r/ollama thread is a community discussion where users share their preferred cloud-based AI models for coding tasks and compare their reasoning capabilities. The conversation likely highlights pop

local-air-ollama
15 Apr 2026
Tutorials

Why Did Apple Fall: Evaluating Curiosity in Large Language Models

DGX agent

arXiv:2510.20635v2 Announce Type: replace-cross Abstract: Curiosity serves as a pivotal conduit for human beings to discover and learn new knowledge. Recent advancements of large language models (LLMs

tutorialsarxiv-cs-ai
15 Apr 2026
Model Releases

BadGraph: A Backdoor Attack Against Latent Diffusion Model for Text-Guided Graph Generation

DGX agent

arXiv:2510.20792v4 Announce Type: replace-cross Abstract: The rapid progress of graph generation has raised new security concerns, particularly regarding backdoor vulnerabilities. While prior work has

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

Belief-Aware VLM Model for Human-like Reasoning

DGX agent

arXiv:2604.09686v1 Announce Type: new Abstract: Traditional neural network models for intent inference rely heavily on observable states and struggle to generalize across diverse tasks and dynamic env

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Can Large Language Models Infer Causal Relationships from Real-World Text?

DGX agent

arXiv:2505.18931v4 Announce Type: replace Abstract: Understanding and inferring causal relationships from texts is a core aspect of human cognition and is essential for advancing large language models

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models

DGX agent

arXiv:2503.21380v3 Announce Type: replace Abstract: The rapid advancement of large reasoning models has saturated existing math benchmarks, underscoring the urgent need for more challenging evaluation

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Computational Lesions in Multilingual Language Models Separate Shared and Language-specific Brain Alignment

DGX agent

arXiv:2604.10627v1 Announce Type: cross Abstract: How the brain supports language across different languages is a basic question in neuroscience and a useful test for multilingual artificial intellige

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Conflicts Make Large Reasoning Models Vulnerable to Attacks

DGX agent

arXiv:2604.09750v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have achieved remarkable performance across diverse domains, yet their decision-making under conflicting objectives rema

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Cross-Cultural Value Awareness in Large Vision-Language Models

DGX agent

arXiv:2604.09945v1 Announce Type: cross Abstract: The rapid adoption of large vision-language models (LVLMs) in recent years has been accompanied by growing fairness concerns due to their propensity t

safetyarxiv-cs-ai
14 Apr 2026
Research

Decoupled Similarity for Task-Aware Token Pruning in Large Vision-Language Models

DGX agent

arXiv:2604.11240v1 Announce Type: new Abstract: Token pruning has emerged as an effective approach to reduce the substantial computational overhead of Large Vision-Language Models (LVLMs) by discardin

researcharxiv-cs-cv
14 Apr 2026
Model Releases

Doc-PP: Document Policy Preservation Benchmark for Large Vision-Language Models

DGX agent

arXiv:2601.03926v2 Announce Type: replace Abstract: The deployment of Large Vision-Language Models (LVLMs) for real-world document question answering is often constrained by dynamic, user-defined poli

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

EgoFun3D: Modeling Interactive Objects from Egocentric Videos using Function Templates

DGX agent

arXiv:2604.11038v1 Announce Type: new Abstract: We present EgoFun3D, a coordinated task formulation, dataset, and benchmark for modeling interactive 3D objects from egocentric videos. Interactive obje

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

Evolutionary Token-Level Prompt Optimization for Diffusion Models

DGX agent

arXiv:2604.09861v1 Announce Type: new Abstract: Text-to-image diffusion models exhibit strong generative performance but remain highly sensitive to prompt formulation, often requiring extensive manual

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

General365: Benchmarking General Reasoning in Large Language Models Across Diverse and Challenging Tasks

DGX agent

arXiv:2604.11778v1 Announce Type: cross Abstract: Contemporary large language models (LLMs) have demonstrated remarkable reasoning capabilities, particularly in specialized domains like mathematics an

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling

DGX agent

arXiv:2604.07209v2 Announce Type: replace Abstract: Building world models with spatial consistency and real-time interactivity remains a fundamental challenge in computer vision. Current video generat

model-releasesarxiv-cs-cv
14 Apr 2026
Tutorials

MorphoFlow: Sparse-Supervised Generative Shape Modeling with Adaptive Latent Relevance

DGX agent

arXiv:2604.11636v1 Announce Type: new Abstract: Statistical shape modeling (SSM) is central to population level analysis of anatomical variability, yet most existing approaches rely on densely annotat

tutorialsarxiv-cs-cv
14 Apr 2026
Industry

New post: We show that small, cheap models can detect the flagship Mythos FreeBSD zero-day (CVE-2026-4747) using a simple harness we call na…

DGX agent

New post: We show that small, cheap models can detect the flagship Mythos FreeBSD zero-day (CVE-2026-4747) using a simple harness we call nano-analyzer Models down to 3.6B active params (including ope

industryclem-delangue--x
14 Apr 2026
Model Releases

NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models: Datasets, Methods and Results

DGX agent

arXiv:2604.10551v1 Announce Type: new Abstract: This paper presents an overview of the NTIRE 2026 Challenge on Short-form UGC Video Restoration in the Wild with Generative Models. This challenge utili

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: AI Flash Portrait (Track 3)

DGX agent

arXiv:2604.11230v1 Announce Type: new Abstract: In this paper, we present a comprehensive overview of the NTIRE 2026 3rd Restore Any Image Model (RAIM) challenge, with a specific focus on Track 3: AI

model-releasesarxiv-cs-cv
14 Apr 2026
Research

Pay Less Attention to Function Words for Free Robustness of Vision-Language Models

DGX agent

arXiv:2512.07222v3 Announce Type: replace-cross Abstract: To address the trade-off between robustness and performance for robust VLM, we observe that function words could incur vulnerability of VLMs a

researcharxiv-cs-cl
14 Apr 2026
Model Releases

RiTeK: A Dataset for Large Language Models Complex Reasoning over Textual Knowledge Graphs in Medicine

DGX agent

arXiv:2410.13987v3 Announce Type: replace Abstract: Answering complex real-world questions in the medical domain often requires accurate retrieval from medical Textual Knowledge Graphs (medical TKGs),

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language Models

DGX agent

arXiv:2505.15489v4 Announce Type: replace-cross Abstract: The impact of multimodal misinformation arises not only from factual inaccuracies but also from the misleading narratives that creators delibe

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Shared Emotion Geometry Across Small Language Models: A Cross-Architecture Study of Representation, Behavior, and Methodological Confounds

DGX agent

arXiv:2604.11050v1 Announce Type: cross Abstract: We extract 21-emotion vector sets from twelve small language models (six architectures x base/instruct, 1B-8B parameters) under a unified comprehensio

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SmileyLlama: Modifying Large Language Models for Directed Chemical Space Exploration

DGX agent

arXiv:2409.02231v5 Announce Type: replace-cross Abstract: We show that large language model (LLMs) can be transformed via supervised fine-tuning (SFT) of engineered prompts into SmileyLlama for explor

model-releasesarxiv-cs-lg
14 Apr 2026
Research

The Weight of a Bit: EMFI Sensitivity Analysis of Embedded Deep Learning Models

DGX agent

arXiv:2602.16309v2 Announce Type: replace-cross Abstract: Fault injection attacks on embedded neural network models have been shown as a potent threat. Numerous works studied resilience of models from

researcharxiv-cs-ai
14 Apr 2026
← Previous
1…119120121122123…1262
Next →