AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,047 results
13 Apr 2026

InstrAct: Towards Action-Centric Understanding in Instructional Videos

SafetyDGX agent

arXiv:2604.08762v1 Announce Type: cross Abstract: Understanding instructional videos requires recognizing fine-grained actions and modeling their temporal relations, which remains challenging for curr

Interactive Program Synthesis for Modeling Collaborative Physical Activities from Narrated Demonstrations

ResearchDGX agent

arXiv:2509.24250v3 Announce Type: replace Abstract: Teaching systems physical tasks is a long standing goal in HCI, yet most prior work has focused on non collaborative physical activities. Collaborat

Large-Scale Universal Defect Generation: Foundation Models and Datasets

ResearchDGX agent

arXiv:2604.08915v1 Announce Type: cross Abstract: Existing defect/anomaly generation methods often rely on few-shot learning, which overfits to specific defect categories due to the lack of large-scal

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LLM4Delay: Flight Delay Prediction via Cross-Modality Adaptation of Large Language Models and Aircraft Trajectory Representation

ResearchDGX agent

arXiv:2510.23636v3 Announce Type: replace-cross Abstract: Flight delay prediction has become a key focus in air traffic management (ATM), as delays reflect inefficiencies in the system. This paper pro

Need help to download from civitai in China

Local AiDGX agent

This Reddit thread from r/StableDiffusion addresses the challenge faced by users in China trying to access and download models from Civitai, which may be restricted or slow due to network limitations

PilotBench: A Benchmark for General Aviation Agents with Safety Constraints

Model ReleasesDGX agent

arXiv:2604.08987v1 Announce Type: new Abstract: As Large Language Models (LLMs) advance toward embodied AI agents operating in physical environments, a fundamental question emerges: can models trained

WOMBET: World Model-based Experience Transfer for Robust and Sample-efficient Reinforcement Learning

TutorialsDGX agent

arXiv:2604.08958v1 Announce Type: cross Abstract: Reinforcement learning (RL) in robotics is often limited by the cost and risk of data collection, motivating experience transfer from a source task to

10 Apr 2026

A Lightweight Library for Energy-Based Joint-Embedding Predictive Architectures

HardwareDGX agent

arXiv:2602.03604v3 Announce Type: replace-cross Abstract: We present EB-JEPA, an open-source library for learning representations and world models using Joint-Embedding Predictive Architectures (JEPAs

A-MBER: Affective Memory Benchmark for Emotion Recognition

Model ReleasesDGX agent

arXiv:2604.07017v1 Announce Type: new Abstract: AI assistants that interact with users over time need to interpret the user's current emotional state in order to respond appropriately and personally.

Ace Step 1.5 XL ComfyUI automation workflow without lama for generating random tags using qwen, generate song and then give it a rating by using waveform analysis

Model ReleasesDGX agent

This is a ComfyUI automation workflow for the ACE-Step 1.5 XL music generation model that uses Qwen (a language model text encoder) to randomly generate music tags/captions without requiring LAMA, ...

Adapting Foundation Models for Annotation-Efficient Adnexal Mass Segmentation in Cine Images

ResearchDGX agent

arXiv:2604.08045v1 Announce Type: new Abstract: Adnexal mass evaluation via ultrasound is a challenging clinical task, often hindered by subjective interpretation and significant inter-observer variab

AgriPath: A Systematic Exploration of Architectural Trade-offs for Crop Disease Classification

Model ReleasesDGX agent

arXiv:2603.13354v3 Announce Type: replace-cross Abstract: Reliable crop disease detection requires models that perform consistently across diverse acquisition conditions, yet existing evaluations ofte

Behavior-Aware Item Modeling via Dynamic Procedural Solution Representations for Knowledge Tracing

ResearchDGX agent

arXiv:2604.08260v1 Announce Type: new Abstract: Knowledge Tracing (KT) aims to predict learners' future performance from past interactions. While recent KT approaches have improved via learning item r

CAAP: Capture-Aware Adversarial Patch Attacks on Palmprint Recognition Models

ResearchDGX agent

arXiv:2604.06987v1 Announce Type: cross Abstract: Palmprint recognition is deployed in security-critical applications, including access control and palm-based payment, due to its contactless acquisiti

ChemVLR: Prioritizing Reasoning in Perception for Chemical Vision-Language Understanding

ApplicationsDGX agent

arXiv:2604.06685v1 Announce Type: cross Abstract: While Vision-Language Models (VLMs) have demonstrated significant potential in chemical visual understanding, current models are predominantly optimiz

Claude Mythos is too dangerous for public consumption...

Model ReleasesDGX agent

Anthropic announced **Claude Mythos Preview**, its most powerful AI model to date, which it is withholding from general public release due to its advanced and potentially dangerous cybersecurity ca...

Consistency-Guided Decoding with Proof-Driven Disambiguation for Three-Way Logical Question Answering

Model ReleasesDGX agent

arXiv:2604.06196v1 Announce Type: cross Abstract: Three-way logical question answering (QA) assigns True/False/Unknown to a hypothesis H given a premise set S. While modern large language models

Evaluating LLMs for Demographic-Targeted Social Bias Detection: A Comprehensive Benchmark Study

Model ReleasesDGX agent

arXiv:2510.04641v3 Announce Type: replace Abstract: Large-scale web-scraped text corpora used to train general-purpose AI models often contain harmful demographic-targeted social biases, creating a re

Extraction of linearized models from pre-trained networks via knowledge distillation

ResearchDGX agent

arXiv:2604.06732v1 Announce Type: new Abstract: Recent developments in hardware, such as photonic integrated circuits and optical devices, are driving demand for research on constructing machine learn

FORGE:Fine-grained Multimodal Evaluation for Manufacturing Scenarios

Model ReleasesDGX agent

arXiv:2604.07413v1 Announce Type: new Abstract: The manufacturing sector is increasingly adopting Multimodal Large Language Models (MLLMs) to transition from simple perception to autonomous execution,

Graph Neural Networks for Misinformation Detection: Performance-Efficiency Trade-offs

Model ReleasesDGX agent

arXiv:2604.08131v1 Announce Type: new Abstract: The rapid spread of online misinformation has led to increasingly complex detection models, including large language models and hybrid architectures. Ho

GroupGPT: A Token-efficient and Privacy-preserving Agentic Framework for Multi-User Chat Assistant

Model ReleasesDGX agent

arXiv:2603.01059v3 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have enabled increasingly capable chatbots. However, most existing systems focus on single-user sett

Inside-Out: Measuring Generalization in Vision Transformers Through Inner Workings

SafetyDGX agent

arXiv:2604.08192v1 Announce Type: cross Abstract: Reliable generalization metrics are fundamental to the evaluation of machine learning models. Especially in high-stakes applications where labeled tar

Knowledge Graphs Generation from Cultural Heritage Texts: Combining LLMs and Ontological Engineering for Scholarly Debates

Model ReleasesDGX agent

arXiv:2511.10354v1 Announce Type: cross Abstract: Cultural Heritage texts contain rich knowledge that is difficult to query systematically due to the challenges of converting unstructured discourse in

Making MLLMs Blind: Adversarial Smuggling Attacks in MLLM Content Moderation

Model ReleasesDGX agent

arXiv:2604.06950v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) are increasingly being deployed as automated content moderators. Within this landscape, we uncover a critic

ModeX: Evaluator-Free Best-of-N Selection for Open-Ended Generation

Model ReleasesDGX agent

arXiv:2601.02535v2 Announce Type: replace Abstract: Selecting a single high-quality output from multiple stochastic generations remains a fundamental challenge for large language models (LLMs), partic

ParkSense: Where Should a Delivery Driver Park? Leveraging Idle AV Compute and Vision-Language Models

AgentsDGX agent

arXiv:2604.07912v1 Announce Type: new Abstract: Finding parking consumes a disproportionate share of food delivery time, yet no system addresses precise parking-spot selection relative to merchant ent

Q-Probe: Scaling Image Quality Assessment to High Resolution via Context-Aware Agentic Probing

Model ReleasesDGX agent

arXiv:2601.15356v4 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) has empowered Multimodal Large Language Models (MLLMs) to achieve superior human preference alignment in Image Qua

Restoring Heterogeneity in LLM-based Social Simulation: An Audience Segmentation Approach

Model ReleasesDGX agent

arXiv:2604.06663v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to simulate social attitudes and behaviors, offering scalable 'silicon samples' that can approximat

Severity-Aware Weighted Loss for Arabic Medical Text Generation

Model ReleasesDGX agent

arXiv:2604.06346v1 Announce Type: cross Abstract: Large language models have shown strong potential for Arabic medical text generation; however, traditional fine-tuning objectives treat all medical ca

Tabular GANs for uneven distribution

Model ReleasesDGX agent

arXiv:2010.00638v2 Announce Type: replace-cross Abstract: Generative models for tabular data have evolved rapidly beyond Generative Adversarial Networks (GANs). While GANs pioneered synthetic tabular

Testimole-Conversational: A 30-Billion-Word Italian Discussion Board Corpus (1996-2024) for Language Modeling and Sociolinguistic Research

ResearchDGX agent

arXiv:2602.14819v2 Announce Type: replace Abstract: We present 'Testimole-conversational' a massive collection of discussion boards messages in the Italian language. The large size of the corpus, more

The AI Skills Shift: Mapping Skill Obsolescence, Emergence, and Transition Pathways in the LLM Era

Model ReleasesDGX agent

arXiv:2604.06906v1 Announce Type: cross Abstract: As Large Language Models reshape the global labor market, policymakers and workers need empirical data on which occupational skills may be most suscep

The Theory and Practice of Highly Scalable Gaussian Process Regression with Nearest Neighbours

Model ReleasesDGX agent

arXiv:2604.07267v1 Announce Type: cross Abstract: Gaussian process (GP) regression is a widely used non-parametric modeling tool, but its cubic complexity in the training size limits its use on mass

ToxReason: A Benchmark for Mechanistic Chemical Toxicity Reasoning via Adverse Outcome Pathway

Model ReleasesDGX agent

arXiv:2604.06264v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have enabled molecular reasoning for property prediction. However, toxicity arises from complex biolog

Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs

Model ReleasesDGX agent

arXiv:2509.08016v2 Announce Type: replace Abstract: Video Large Language Models (VideoLLMs) face a critical bottleneck: increasing the number of input frames to capture fine-grained temporal detail le

WASD: Locating Critical Neurons as Sufficient Conditions for Explaining and Controlling LLM Behavior

Model ReleasesDGX agent

arXiv:2603.18474v2 Announce Type: replace Abstract: Precise behavioral control of large language models (LLMs) is critical for complex applications. However, existing methods often incur high training

9 Apr 2026

Claude Mythos and misguided open-weight fearmongering

Model ReleasesDGX agent

The *Interconnects.ai* article argues that fears around releasing an open-weight version of Claude Mythos are overstated, noting that closed frontier models still lead open-weight ones in robust, ...

8 Apr 2026

Google just casually disrupted the open-source AI narrative…

TutorialsDGX agent

Google released Gemma 4 on April 2, 2026 — a family of four open-weight models built on the same research as Gemini 3 and licensed under the permissive Apache 2.0 license, marking a significant shi...

7 Apr 2026

Check out the GLM-5.1 first impressions with Peter on our YouTube https://www.youtube.com/watch?v=f11tVBXWr2g

Model ReleasesDGX agent

Z.ai's GLM-5.1 is a 754-billion parameter open-weight Mixture-of-Experts model released on April 7, 2026 under an MIT license, designed as a post-training upgrade to GLM-5 with a focus on long-hori...

GLM-5.1 is now available in Go w/ Zero Data Retention

Model ReleasesDGX agent

GLM-5.1 is now available through OpenCode Go, a low-cost subscription service ($5 first month, then $10/month) that provides access to open-source coding models with a zero data retention policy, m...

GLM-5.1: Towards Long-Horizon Tasks

Model ReleasesDGX agent

GLM-5.1: Towards Long-Horizon Tasks Chinese AI lab Z.ai's latest model is a giant 754B parameter 1.51TB (on Hugging Face) MIT-licensed monster - the same size as their previous GLM-5 release, and shar

Ok @cognition SWE-1.6 Fast is better than expected. Had it built a quick prototype UI, looks better than what Figma Make generated and is ri…

AgentsDGX agent

A user on X (@willebrew) shared positive impressions of Cognition's SWE-1.6 Fast model, noting that a prototype UI it generated surpassed the output of Figma Make. SWE-1.6 is Cognition's latest mo...

17 Aug 2026

A Year in LLM Serving: Workload Evolution, Caching and Load-Balancing

ApplicationsDGX agent

arXiv:2608.13573v1 Announce Type: new Abstract: Large Language Model (LLM) serving has become a critical cloud workload, and realistic traces are essential for motivating and benchmarking serving syst

After pushing 1M+ tokens through Qwen 3.8 27B, here is my optimal llama.cpp config for 16GB VRAM (73k Context, Agentic Coding)

Model ReleasesDGX agent

Following up on my previous post about my budget server setup (Intel N100 + RTX 5060 Ti 16GB), a few of you asked for a deeper dive into my actual inference config and real-world agentic performance.

AnchorBench: A Multi-Pathway Benchmark for the Anchoring Effect in LLMs

Model ReleasesDGX agent

arXiv:2608.14320v1 Announce Type: new Abstract: The anchoring effect is a cognitive bias in which an initial reference value shifts a later judgment toward itself. This effect is well established in h

Automated Inference of Graph Transformation Rules

ResearchDGX agent

arXiv:2404.02692v3 Announce Type: replace-cross Abstract: The explosion of data available in life sciences is fueling an increasing demand for expressive models and computational methods. Graph transf

Bootstrapping Niche Multilingual Code Translation via Reinforcement Learning with Execution-Based Verifiable Supervision

Model ReleasesDGX agent

arXiv:2608.13854v1 Announce Type: new Abstract: Code translation must preserve executable behavior across many programming languages, yet neural code translation has largely focused on a few popular l

ChartProbe: A Diagnostic Study on Visual Reasoning through Perception, Grounding, and Simple Reasoning

Model ReleasesDGX agent

arXiv:2608.13766v1 Announce Type: new Abstract: Vision-language models (VLMs) remain unreliable on chart questions that require reasoning over visual quantities, and this weakness is usually attribute

Concept Guidance: Precise, Training-Free Latent Control for Text-to-Image Generation

ResearchDGX agent

arXiv:2608.14172v1 Announce Type: cross Abstract: Text-to-image diffusion models have two major drawbacks that severely limit their practical utility: (1) standard models lack an intrinsic mechanism f

CoViLLM: An Adaptive Human-Robot Collaborative Assembly Framework Using Large Language Models

Local AiDGX agent

arXiv:2603.11461v3 Announce Type: replace Abstract: With increasing demand for mass customization, traditional manufacturing robots that rely on rule-based operations lack the flexibility to accommoda

Don't Claim Benchmark-Oriented Optimization Improves General Coding Capability -- Diverse Evaluation Is Required

Model ReleasesDGX agent

arXiv:2608.13566v1 Announce Type: cross Abstract: Post-training papers, model cards, and blog posts often treat scores on a small set of coding benchmarks (e.g., SWE-bench and LiveCodeBench) as eviden

IIRC, this is the 3rd month in a row that @FireworksAI_HQ is being featured on @tryramp's top software vendor list. This isn't just for infe…

Model ReleasesDGX agent

IIRC, this is the 3rd month in a row that @FireworksAI_HQ is being featured on @tryramp's top software vendor list. This isn't just for inference / model serving (of which we serve 40T+ tokens a day),

Nanbeige4.2-3B on Apple Silicon: Fixing Deployment Bugs and Decreasing Looped Transformer Memory Overhead

Model ReleasesDGX agent

arXiv:2608.13987v1 Announce Type: new Abstract: Nanbeige4.2-3B is a 3B-parameter agentic model built around a Looped Transformer (LT) that reuses one stack of layers for a second forward pass, adding

Offline Deep Q* Estimation with Diffusion Models

ResearchDGX agent

arXiv:2608.14401v1 Announce Type: cross Abstract: In offline RL, estimating the optimal action-value function Q^* can be formulated as solving the optimal Bellman equation based solely on offline obse

Ontology-Grounded World Models for Failure Diagnosis and Closed-Loop Repair in Physical AI Systems

ResearchDGX agent

arXiv:2608.13901v1 Announce Type: new Abstract: EV-WM represents candidate quality with feature and event scores, but these scores do not explicitly record an unmet task predicate, a route label for a

P2Skill: Privacy Preserving Skill Distillation for Cloud-Local LLM Inference Systems

Model ReleasesDGX agent

arXiv:2608.14094v1 Announce Type: cross Abstract: Cloud-local LLM inference systems have the potential to use the reasoning capability of large cloud models while protecting sensitive user data on per

QuaSAR: Quantization Compensation via Stable Activation-Aware Rank Truncation

Model ReleasesDGX agent

arXiv:2608.14149v1 Announce Type: new Abstract: Recent training-free post-training quantization methods restore model accuracy through closed-form residual compensation. To constrain additional model

Qwen3.8-27B Q8_0 on Strix Halo is seriously impressive

Model ReleasesDGX agent

Sorry for the slop, but I was impressed by this model as I have been testing Qwen3.8-27B Q8_0 locally on my ROG Flow Z13 (Ryzen AI Max+ 395, 128 GB unified memory) and this model was the only one who

Reading Between The Lines: Modeling and Evaluating Behavioral Realism in Legal Simulation

ApplicationsDGX agent

arXiv:2608.13712v1 Announce Type: cross Abstract: Deposition training requires attorneys to manage dynamic witness behavior, yet legal-AI evaluations largely focus on factual accuracy, reasoning, or r

← Previous
1…242243244245246…1018
Next →