AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlog
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,172 results
Research

Impact of Nonlinear Power Amplifier on Massive MIMO: Machine Learning Prediction Under Realistic Radio Channel

DGX agent

arXiv:2604.15977v1 Announce Type: new Abstract: M-MIMO is one of the crucial technologies for increasing spectral and energy efficiency of wireless networks. Most of the current works assume that M-MI

researcharxiv-cs-lg
20 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

InstructTable: Improving Table Structure Recognition Through Instructions

DGX agent

arXiv:2604.02880v2 Announce Type: replace Abstract: Table structure recognition (TSR) holds widespread practical importance by parsing tabular images into structured representations, yet encounters si

model-releasesarxiv-cs-cv
20 Apr 2026
Industry

Kimi-K2.6 is on HuggingFace

DGX agent

Kimi-K2.6 is a language model that has been released on HuggingFace, a popular platform for sharing machine learning models and datasets. The announcement was made by Clem Delangue, likely indicating

industryclem-delangue--x
20 Apr 2026
Model Releases

LLMSniffer: Detecting LLM-Generated Code via GraphCodeBERT and Supervised Contrastive Learning

DGX agent

arXiv:2604.16058v1 Announce Type: cross Abstract: The rapid proliferation of Large Language Models (LLMs) in software development has made distinguishing AI-generated code from human-written code a cr

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

MemEvoBench: Benchmarking Memory MisEvolution in LLM Agents

DGX agent

arXiv:2604.15774v1 Announce Type: new Abstract: Equipping Large Language Models (LLMs) with persistent memory enhances interaction continuity and personalization but introduces new safety risks. Speci

model-releasesarxiv-cs-cl
20 Apr 2026
Research

Neural Continuous-Time Markov Chain: Discrete Diffusion via Decoupled Jump Timing and Direction

DGX agent

arXiv:2604.15694v1 Announce Type: new Abstract: Discrete diffusion models based on continuous-time Markov chains (CTMCs) have shown strong performance on language and discrete data generation, yet exi

researcharxiv-cs-lg
20 Apr 2026
Model Releases

Ragged Paged Attention: A High-Performance and Flexible LLM Inference Kernel for TPU

DGX agent

arXiv:2604.15464v1 Announce Type: cross Abstract: Large Language Model (LLM) deployment is increasingly shifting to cost-efficient accelerators like Google's Tensor Processing Units (TPUs), prioritizi

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Scalable Maximum Entropy Population Synthesis via Persistent Contrastive Divergence

DGX agent

arXiv:2603.27312v2 Announce Type: replace Abstract: Maximum entropy (MaxEnt) modelling provides a principled framework for generating synthetic populations from aggregate census data, without access t

model-releasesarxiv-cs-lg
20 Apr 2026
Model Releases

Social-JEPA: Emergent Geometric Isomorphism

DGX agent

arXiv:2603.02263v2 Announce Type: replace-cross Abstract: World models compress rich sensory streams into compact latent codes that anticipate future observations. We let separate agents acquire such

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

SocialGrid: A Benchmark for Planning and Social Reasoning in Embodied Multi-Agent Systems

DGX agent

arXiv:2604.16022v1 Announce Type: new Abstract: As Large Language Models (LLMs) transition from text processors to autonomous agents, evaluating their social reasoning in embodied multi-agent settings

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Solving Inverse Parametrized Problems via Finite Elements and Extreme Learning Networks

DGX agent

arXiv:2602.14757v2 Announce Type: replace-cross Abstract: We develop an interpolation-based modeling framework for parameter-dependent partial differential equations arising in control, inverse proble

model-releasesarxiv-cs-lg
20 Apr 2026
Safety

Targeted Exploration via Unified Entropy Control for Reinforcement Learning

DGX agent

arXiv:2604.14646v2 Announce Type: replace Abstract: Recent advances in reinforcement learning (RL) have improved the reasoning capabilities of large language models (LLMs) and vision-language models (

safetyarxiv-cs-ai
20 Apr 2026
Research

The Amazing Stability of Flow Matching

DGX agent

arXiv:2604.16079v1 Announce Type: new Abstract: The success of deep generative models in generating high-quality and diverse samples is often attributed to particular architectures and large training

researcharxiv-cs-cv
20 Apr 2026
Model Releases

The Jensen + @dwarkesh_sp podcast was fantastic. Jensen is someone who understood how ecosystems work and someone who understands real-world…

DGX agent

The Jensen + @dwarkesh_sp podcast was fantastic. Jensen is someone who understood how ecosystems work and someone who understands real-world trade, policy and controls work. And in some deeper sense h

model-releasessoumith-chintala--x
20 Apr 2026
Model Releases

VEFX-Bench: A Holistic Benchmark for Generic Video Editing and Visual Effects

DGX agent

arXiv:2604.16272v1 Announce Type: cross Abstract: As AI-assisted video creation becomes increasingly practical, instruction-guided video editing has become essential for refining generated or captured

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

When Cultures Meet: Multicultural Text-to-Image Generation

DGX agent

arXiv:2502.15972v2 Announce Type: replace-cross Abstract: Text-to-image generation models have achieved strong performance in culturally homogeneous settings, yet their ability to generate multicultur

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Where does output diversity collapse in post-training?

DGX agent

arXiv:2604.16027v1 Announce Type: cross Abstract: Post-trained language models produce less varied outputs than their base counterparts. This output diversity collapse undermines inference-time scalin

researcharxiv-cs-ai
20 Apr 2026
Model Releases

Wisdom is Knowing What not to Say: Hallucination-Free LLMs Unlearning via Attention Shifting

DGX agent

arXiv:2510.17210v3 Announce Type: replace Abstract: The increase in computing power and the necessity of AI-assisted decision-making boost the growing application of large language models (LLMs). Alon

model-releasesarxiv-cs-cl
20 Apr 2026
Local Ai

在 Ollama 运行 hermes

DGX agent

This post likely discusses how to run the Hermes language model using Ollama, an open-source tool for running large language models locally. It probably provides instructions or insights on setting up

local-aiollama--x
18 Apr 2026
Research

A Linguistics-Aware LLM Watermarking via Syntactic Predictability

DGX agent

arXiv:2510.13829v3 Announce Type: replace Abstract: As large language models (LLMs) continue to advance rapidly, reliable governance tools have become critical. Publicly verifiable watermarking is par

researcharxiv-cs-cl
17 Apr 2026
Model Releases

AccelOpt: A Self-Improving LLM Agentic System for AI Accelerator Kernel Optimization

DGX agent

arXiv:2511.15915v2 Announce Type: replace-cross Abstract: We present AccelOpt, a self-improving large language model (LLM) agentic system that autonomously optimizes kernels for emerging AI acclerator

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Anomaly Detection in IEC-61850 GOOSE Networks: Evaluating Unsupervised and Temporal Learning for Real-Time Intrusion Detection

DGX agent

arXiv:2604.14233v1 Announce Type: cross Abstract: The IEC-61850 GOOSE protocol underpins time-critical communication in modern digital substations but lacks native security mechanisms, leaving it vuln

researcharxiv-cs-lg
17 Apr 2026
Model Releases

Benchmarking Optimizers for MLPs in Tabular Deep Learning

DGX agent

arXiv:2604.15297v1 Announce Type: new Abstract: MLP is a heavily used backbone in modern deep learning (DL) architectures for supervised learning on tabular data, and AdamW is the go-to optimizer used

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

CaptionQA: Is Your Caption as Useful as the Image Itself?

DGX agent

arXiv:2511.21025v2 Announce Type: replace Abstract: Image captions serve as efficient surrogates for visual content in multimodal systems such as retrieval, recommendation, and multi-step agentic infe

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

Chinese Language Is Not More Efficient Than English in Vibe Coding: A Preliminary Study on Token Cost and Problem-Solving Rate

DGX agent

arXiv:2604.14210v1 Announce Type: new Abstract: A claim has been circulating on social media and practitioner forums that Chinese prompts are more token-efficient than English for LLM coding tasks, po

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Claude Opus 4.7 on AI Gateway

DGX agent

Claude Opus 4.7 is now available through Vercel's AI Gateway, Anthropic's latest large language model offering integration with Vercel's platform for developers. This enables developers to access Clau

model-releasesvercel-blog
17 Apr 2026
Safety

Context Over Content: Exposing Evaluation Faking in Automated Judges

DGX agent

arXiv:2604.15224v1 Announce Type: cross Abstract: The extit{LLM-as-a-judge} paradigm has become the operational backbone of automated AI evaluation pipelines, yet rests on an unverified assumption: th

safetyarxiv-cs-cl
17 Apr 2026
Research

Data Synthesis Improves 3D Myotube Instance Segmentation

DGX agent

arXiv:2604.14720v1 Announce Type: new Abstract: Myotubes are multinucleated muscle fibers serving as key model systems for studying muscle physiology, disease mechanisms, and drug responses. Mechanist

researcharxiv-cs-cv
17 Apr 2026
Research

Faithfulness Serum: Mitigating the Faithfulness Gap in Textual Explanations of LLM Decisions via Attribution Guidance

DGX agent

arXiv:2604.14325v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong performance and have revolutionized NLP, but their lack of explainability keeps them treated as black boxes,

researcharxiv-cs-cl
17 Apr 2026
Model Releases

GeoAgentBench: A Dynamic Execution Benchmark for Tool-Augmented Agents in Spatial Analysis

DGX agent

arXiv:2604.13888v1 Announce Type: new Abstract: The integration of Large Language Models (LLMs) into Geographic Information Systems (GIS) marks a paradigm shift toward autonomous spatial analysis. How

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

I have found 4.7 great for design, reverted back to 4.6 extended for everything else Anyone else like this?

DGX agent

I have found 4.7 great for design, reverted back to 4.6 extended for everything else Anyone else like this? Introducing Claude Design by Anthropic Labs: make prototypes, slides, and one-pagers by talk

model-releasesemad-mostaque--x
17 Apr 2026
Model Releases

Knowing When Not to Answer: Evaluating Abstention in Multimodal Reasoning Systems

DGX agent

arXiv:2604.14799v1 Announce Type: new Abstract: Effective abstention (EA), recognizing evidence insufficiency and refraining from answering, is critical for reliable multimodal systems. Yet existing e

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

MADE: A Living Benchmark for Multi-Label Text Classification with Uncertainty Quantification of Medical Device Adverse Events

DGX agent

arXiv:2604.15203v1 Announce Type: new Abstract: Machine learning in high-stakes domains such as healthcare requires not only strong predictive performance but also reliable uncertainty quantification

model-releasesarxiv-cs-cl
17 Apr 2026
Local Ai

One RL to See Them All: Visual Triple Unified Reinforcement Learning

DGX agent

arXiv:2505.18129v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) is becoming an important direction for post-training vision-language models (VLMs), but public training methodolog

local-aiarxiv-cs-cl
17 Apr 2026
Model Releases

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games

DGX agent

arXiv:2506.03610v3 Announce Type: replace Abstract: Large Language Model (LLM) agents are reshaping the game industry, by enabling more intelligent and human-preferable characters. Yet, current game b

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

PeerPrism: Peer Evaluation Expertise vs Review-writing AI

DGX agent

arXiv:2604.14513v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used in scientific peer review, assisting with drafting, rewriting, expansion, and refinement. However, ex

model-releasesarxiv-cs-cl
17 Apr 2026
Research

PixelDiT: Pixel Diffusion Transformers for Image Generation

DGX agent

arXiv:2511.20645v2 Announce Type: replace Abstract: Latent-space modeling has been the standard for Diffusion Transformers (DiTs). However, it relies on a two-stage pipeline where the pretrained autoe

researcharxiv-cs-cv
17 Apr 2026
Safety

Pushing the Boundaries of Multiple Choice Evaluation to One Hundred Options

DGX agent

arXiv:2604.14634v1 Announce Type: new Abstract: Multiple choice evaluation is widely used for benchmarking large language models, yet near ceiling accuracy in low option settings can be sustained by s

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

SAGE Celer 2.6 Technical Card

DGX agent

arXiv:2604.14168v1 Announce Type: new Abstract: We introduce SAGE Celer 2.6, the latest in our line of general-purpose Celer models from SAGEA. Celer 2.6 is available in 5B, 10B, and 27B parameter siz

model-releasesarxiv-cs-cl
17 Apr 2026
Applications

Standard-to-Dialect Transfer Trends Differ across Text and Speech: A Case Study on Intent and Topic Classification in German Dialects

DGX agent

arXiv:2510.07890v3 Announce Type: replace Abstract: Research on cross-dialectal transfer from a standard to a non-standard dialect variety has typically focused on text data. However, dialects are pri

applicationsarxiv-cs-cl
17 Apr 2026
Local Ai

The Courtroom Trial of Pixels: Robust Image Manipulation Localization via Adversarial Evidence and Reinforcement Learning Judgment

DGX agent

arXiv:2604.14703v1 Announce Type: new Abstract: Although some existing image manipulation localization (IML) methods incorporate authenticity-related supervision, this information is typically utilize

local-aiarxiv-cs-cv
17 Apr 2026
Safety

To See or To Please: Uncovering Visual Sycophancy and Split Beliefs in VLMs

DGX agent

arXiv:2603.18373v2 Announce Type: replace Abstract: When VLMs answer correctly, do they genuinely rely on visual information or exploit language shortcuts? We introduce the Tri-Layer Diagnostic Framew

safetyarxiv-cs-cv
17 Apr 2026
Model Releases

VeruSAGE: A Study of Agent-Based Verification for Rust Systems

DGX agent

arXiv:2512.18436v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown impressive capability to understand and develop code. However, their capability to rigorously reason a

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

xFODE+: Explainable Type-2 Fuzzy Additive ODEs for Uncertainty Quantification

DGX agent

arXiv:2604.14880v1 Announce Type: new Abstract: Recent advances in Deep Learning (DL) have boosted data-driven System Identification (SysID), but reliable use requires Uncertainty Quantification (UQ)

model-releasesarxiv-cs-lg
17 Apr 2026
Safety

Your LLM Agents are Temporally Blind: The Misalignment Between Tool Use Decisions and Human Time Perception

DGX agent

arXiv:2510.23853v3 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly used to interact with and execute tasks in dynamic environments. However, a critical yet overlook

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

An Optimal Transport-driven Approach for Cultivating Latent Space in Online Incremental Learning

DGX agent

arXiv:2211.16780v3 Announce Type: replace-cross Abstract: In online incremental learning, data continuously arrives with substantial distributional shifts, creating a significant challenge because pre

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Anthropic launches Claude Opus 4.7 with coding, visual reasoning improvements

DGX agent

Anthropic PBC today opened access to Claude Opus 4.7, the latest addition to its popular line of large language models. The company says that the LLM is significantly better than its predecessor at co

model-releasessiliconangle
16 Apr 2026
Local Ai

Best Ollama models/settings for an 8GB VPS (CPU only, ARM)? Running into memory & looping issues.

DGX agent

This Reddit thread discusses running Ollama on a resource-constrained 8GB CPU-only ARM VPS, addressing common challenges such as out-of-memory errors and model response looping. For purely CPU-only se

local-air-ollama
16 Apr 2026
← Previous
1…476477478479480…1358
Next →