AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,620 results
Model Releases

VisReason: A Large-Scale Dataset for Visual Chain-of-Thought Reasoning

DGX agent

arXiv:2511.17731v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting has proven remarkably effective for eliciting complex reasoning in large language models (LLMs). Yet, its potential

model-releasesarxiv-cs-cv
2 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

VolumeDP: Modeling Volumetric Representation for Manipulation Policy Learning

DGX agent

arXiv:2603.17720v2 Announce Type: replace Abstract: Imitation learning is a prominent paradigm for robotic manipulation. However, existing visual imitation methods map 2D image observations directly t

model-releasesarxiv-cs-ro
2 Jul 2026
Model Releases

Wake up for Touch! Mask-isolated Tactile Alignment Learning in MLLMs

DGX agent

arXiv:2607.00302v1 Announce Type: new Abstract: Touch supplies the physical grounding needed to perceive intrinsic material properties, such as friction and compliance, that vision alone often cannot

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

We are hiring our founding team in Korea 🇰🇷 Join us! P.S. Mistral will be at @icmlconf (July 6–11). Come meet the team!

DGX agent

Mistral AI is recruiting for its founding team in Korea and will have representatives attending ICML conference from July 6-11, 2024, where interested candidates can meet the team in person.

model-releasesarthur-mensch--x
2 Jul 2026
Model Releases

What's Hidden Matters: Identifying Planning-Critical Occluded Agents using Vision-Language Models

DGX agent

arXiv:2607.00283v1 Announce Type: cross Abstract: Autonomous vehicles must safely navigate complex environments where planning-critical agents may be hidden from view. Current approaches often treat a

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Why Advanced Encoders Lag on Sparse Retrieval? The Answer and an Approach to Bridging Vocabulary Gaps

DGX agent

arXiv:2607.00004v1 Announce Type: cross Abstract: While advanced foundation models like ModernBERT significantly outperform older architectures in dense retrieval, they surprisingly lag behind the agi

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Wordle 1,839 4/6 ⬛⬛🟨⬛🟨 ⬛⬛🟨⬛⬛ ⬛🟨⬛🟨🟩 🟩🟩🟩🟩🟩

DGX agent

I cannot provide a meaningful summary for this entry as the content appears to be a personal Wordle game result (puzzle #1,839 solved in 4 attempts) rather than substantive knowledge base material. Th

model-releasesanthropic--x
2 Jul 2026
Model Releases

WorkBench Revisited: Workplace Agents Two Years On

DGX agent

arXiv:2606.13715v2 Announce Type: replace Abstract: The best agent on WorkBench in March 2024, GPT-4, completed just 43% of tasks. We revisit the benchmark in June 2026 and find that the best agent to

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

XSkill: Continual Learning from Experience and Skills in Multimodal Agents

DGX agent

arXiv:2603.12056v3 Announce Type: replace Abstract: Multimodal agents can now tackle complex reasoning tasks with diverse tools, yet they still suffer from inefficient tool use and inflexible orchestr

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

YOMI-Bench: A Benchmark for Evaluating Kanji Reading and Phonological Understanding of LLMs for Japanese

DGX agent

arXiv:2607.00664v1 Announce Type: new Abstract: We propose YOMI-Bench, a benchmark for evaluating kanji reading and phonological understanding of large language models (LLMs) for Japanese. In Japanese

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

You really need your own benchmarks. If you are translating hieroglyphics, use Gemini 3.5 Flash. If you are running a vending machine use Op…

DGX agent

You really need your own benchmarks. If you are translating hieroglyphics, use Gemini 3.5 Flash. If you are running a vending machine use Opus 4.8. (This is one reason why I am skeptical of just swapp

model-releasesethan-mollick--x
2 Jul 2026
Model Releases

Your coding agent bill doubled and nobody can tell you why. Here's the actual reason: Claude Code, Cursor, and Copilot all log activity in d…

DGX agent

Your coding agent bill doubled and nobody can tell you why. Here's the actual reason: Claude Code, Cursor, and Copilot all log activity in different formats. The second your team uses more than one (t

model-releasesharrison-chase--x
2 Jul 2026
Model Releases

Z.ai launches ZCode, an 'Agentic Development Environment' optimized for its new GLM-5.2 model; Z.ai's GLM Coding Plan costs from 16.20 to 144 per month (Michael Nuñez/VentureBeat)

DGX agent

Michael Nuñez / VentureBeat: Z.ai launches ZCode, an “Agentic Development Environment” optimized for its new GLM-5.2 model; Z.ai's GLM Coding Plan costs from 16.20 to 144 per month — The move marks th

model-releasestechmeme
2 Jul 2026
Model Releases

ZO-Act: Efficient Zeroth-Order Fine-Tuning via One-Shot Activation-Informed Low-Rank Subspaces

DGX agent

arXiv:2607.01125v1 Announce Type: new Abstract: Zeroth-order (ZO) optimization enables fine-tuning large language models when backpropagation is unavailable or memory-prohibitive, but existing methods

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

A Large-Language-Model Supported Personalized Driving Framework for Lane Change in Highway Scenarios

DGX agent

arXiv:2606.31483v1 Announce Type: new Abstract: Personalized driving can improve the user acceptance of automated driving systems. However, existing methods still provide limited support for translati

model-releasesarxiv-cs-ro
1 Jul 2026
Model Releases

A Realistic Protocol for Evaluation of Weakly Supervised Object Localization

DGX agent

arXiv:2404.10034v3 Announce Type: replace Abstract: Weakly Supervised Object Localization (WSOL) allows training deep learning models for classification and localization (LOC) using only global class-

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

A Reproducible Benchmark of Lightweight CNNs: Accuracy, Efficiency, and the Impact of Pretrained Initialization

DGX agent

arXiv:2505.03303v3 Announce Type: replace-cross Abstract: Lightweight convolutional neural networks are often compared using results obtained with different training recipes, input settings, and pretr

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

A Self-Evolving Agentic System for Automated Generation and Execution of Biological Protocols

DGX agent

arXiv:2606.31763v1 Announce Type: new Abstract: Autonomous wet-lab experimentation requires more than plausible protocol text: biological intent, quantitative procedures, device constraints and experi

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

A Semantic-Layer-Mediated Agent for Natural Language to SQL over Heterogeneous Enterprise Databases

DGX agent

arXiv:2606.31041v1 Announce Type: new Abstract: Natural language-to-SQL (NL2SQL) over real-world enterprise databases remains significantly more challenging than on academic benchmarks. Enterprise sch

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

A Stationary-Distribution Theory for Triplet-Based Plateau Search in Random Forest Ensemble-Size Selection

DGX agent

arXiv:2606.30837v1 Announce Type: cross Abstract: The number of trees is a central computational parameter in Random Forests: increasing it reduces finite-ensemble variability but increases training a

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

A swap-adversarial framework for improving domain generalization in electrocorticography-based Parkinson's disease classification

DGX agent

arXiv:2602.10528v2 Announce Type: replace-cross Abstract: We propose a novel swap-adversarial framework that mitigates high inter-subject variability and the high-dimensional low-sample-size problem i

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

A Systematic Approach to Multi-Agent AI from Advanced Regulatory Control Theory: Safe and Auditable LLM Operator Agents for Process Control

DGX agent

arXiv:2606.30877v1 Announce Type: cross Abstract: Recent literature shows that large language models (LLMs) are useful for general-purpose tasks yet perform poorly on specific domain ones. One reason

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

A Three-Phase Foundation Model for Tax-Aware Personalized Portfolio Management

DGX agent

arXiv:2606.30997v1 Announce Type: new Abstract: We present a three-phase deep reinforcement learning system for personalized portfolio management that addresses three limitations shared by all prior f

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

A time-series classification framework for individual-level absenteeism prediction under severe class imbalance

DGX agent

arXiv:2606.31532v1 Announce Type: new Abstract: Staff absenteeism imposes substantial operational costs in high-demand work environments such as healthcare, emergency services, meat processing, constr

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

A Transferable Learned Temporal Prior for Transmission Reconstruction and Decision-Relevant Uncertainty in Real Outbreak Labels

DGX agent

arXiv:2606.30842v1 Announce Type: new Abstract: Outbreak transmission reconstruction treats epidemiological timing and transmission labels as deterministic ground truth; neither has been systematicall

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

Absorption-Feature-Guided Distance-Decoupled Estimation and Band Selection for LWIR Hyperspectral Passive Ranging

DGX agent

arXiv:2606.31824v1 Announce Type: new Abstract: Long-wave infrared (LWIR) hyperspectral observations contain distance-dependent atmospheric absorption signatures, providing a physical basis for long-r

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

Accelerometry-Derived Digital Biomarkers for Cardiometabolic Risk: A Population-Representative Tabular Benchmark with Uncertainty Quantification

DGX agent

arXiv:2606.30702v1 Announce Type: cross Abstract: Structured tabular data dominates clinical medicine, yet existing benchmarks fail to reflect real-world properties like complex survey sampling, demog

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Adaptive Cluster-First Route-Second Decomposition for Industrial-Scale Vehicle Routing

DGX agent

arXiv:2606.31820v1 Announce Type: new Abstract: Large-scale capacitated vehicle routing problems (CVRPs) are commonly addressed using cluster-first route-second (CFRS) approaches that split a routing

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Again, 2 weeks missed. No rehab. Rib injury, which really affects ability to rotate (comfortably) and he's just raking. He's on a heater.

DGX agent

Again, 2 weeks missed. No rehab. Rib injury, which really affects ability to rotate (comfortably) and he's just raking. He's on a heater. It is kind of unfathomable that Chase DeLauter missed two week

model-releasesanthropic--x
1 Jul 2026
Model Releases

AgentBound: Verifiable Behavioral Governance for Autonomous AI Agents

DGX agent

arXiv:2606.30970v1 Announce Type: new Abstract: Autonomous AI agents increasingly perform consequential actions on behalf of human principals, including financial transactions, external communications

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

“Agentic kernel optimization is the future of on-device inference” @xenovacom used Fable 5 to write kernels that pushed Gemma 4 to a massive…

DGX agent

“Agentic kernel optimization is the future of on-device inference” @xenovacom used Fable 5 to write kernels that pushed Gemma 4 to a massive 255 tok/s on WebGPU with M4. He shared the demo, so you can

model-releasesclem-delangue--x
1 Jul 2026
Model Releases

Agentic RAG-VLM: Affordance-Aware Retrieval-Augmented Generation with Self-Reflective Planning for Robotic Grasping

DGX agent

arXiv:2606.31200v1 Announce Type: new Abstract: Generalizable robotic grasping in cluttered environments is essential for deploying manipulators in unstructured human spaces, yet existing VLM-based me

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Agentic retrieval is changing the way retrieval-augmented applications are built, especially in domains like legal and fintech, where agents…

DGX agent

Agentic retrieval is changing the way retrieval-augmented applications are built, especially in domains like legal and fintech, where agents need to autonomously navigate large, evolving knowledge bas

model-releasesllamaindex--x
1 Jul 2026
Model Releases

AION: Aerial Indoor Object-Goal Navigation Using Dual-Policy Reinforcement Learning

DGX agent

arXiv:2601.15614v3 Announce Type: replace Abstract: Object-Goal Navigation (ObjectNav) requires an agent to autonomously explore an unknown environment and navigate toward target objects specified by

model-releasesarxiv-cs-ro
1 Jul 2026
Model Releases

AlloyDB AI Functions - now with revolutionary performance boosts and cost savings

DGX agent

AlloyDB is an AI-native database—it isn’t just a passive data store, it intelligently understands and processes your data. With AlloyDB, you get industry-leading vector and hybrid search, near 100% ac

model-releasesgoogle-cloud-ai
1 Jul 2026
Model Releases

An Empirical Study of Security Calibration in Large Language Models for Code

DGX agent

arXiv:2606.31159v1 Announce Type: cross Abstract: Large Language Models (LLMs) are rapidly transforming software development, yet their use in security-critical contexts raises a key question: do mode

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

An Executable Benchmarking Suite for Tool-Using Agents

DGX agent

arXiv:2605.11030v2 Announce Type: replace-cross Abstract: Closed-loop tool-using agents are increasingly evaluated in executable web, code, and micro-task environments, but benchmark reports often con

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Anthropic launches Claude Sonnet 5 AI model with coding, safety upgrades as Fable and Mythos controls lifted

DGX agent

Anthropic PBC today debuted Claude Sonnet 5, a midrange large language model that outperforms its predecessor in several areas. The LLM will be the default option in the consumer tiers of the company’

model-releasessiliconangle
1 Jul 2026
Model Releases

Anthropic says Fable 5 will be available via usage credits for Claude users from July 7, and is working with partners to draft an AI jailbreak severity standard (Anthropic)

DGX agent

Anthropic: Anthropic says Fable 5 will be available via usage credits for Claude users from July 7, and is working with partners to draft an AI jailbreak severity standard — On Friday, June 12, the US

model-releasestechmeme
1 Jul 2026
Model Releases

Anthropic says it is rolling back a covert Claude Code tracking feature to identify users based in China or affiliated with Chinese AI labs, after backlash (Juro Osawa/The Information)

DGX agent

Juro Osawa / The Information: Anthropic says it is rolling back a covert Claude Code tracking feature to identify users based in China or affiliated with Chinese AI labs, after backlash — Anthropic is

model-releasestechmeme
1 Jul 2026
Model Releases

Anthropic says 'some routine tasks like coding and debugging' on Fable 5 'will fall back to Opus 4.8' in 'the near term' as it works to 'reduce false positives' (@anthropicai)

DGX agent

@anthropicai: Anthropic says “some routine tasks like coding and debugging” on Fable 5 “will fall back to Opus 4.8” in “the near term” as it works to “reduce false positives” — Claude Fable 5 will be

model-releasestechmeme
1 Jul 2026
Model Releases

ARC-AGI-3 is built different, it has dumbfounded almost all regular attempts so far because it's so much harder than anything that came befo…

DGX agent

ARC-AGI-3 is built different, it has dumbfounded almost all regular attempts so far because it's so much harder than anything that came before. It has no rules, it's agentic and has no explicit goals,

model-releasesfrancois-chollet--x
1 Jul 2026
Model Releases

Arena-T2I Hard: Benchmarking and Improving Faithfulness with Dependency-Aware Checklist

DGX agent

arXiv:2606.31711v1 Announce Type: new Abstract: Faithfulness -- how precisely a generated image aligns with its prompt -- is increasingly central to the real-world utility of text-to-image (T2I) model

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Armadin details full sandbox escape in Claude Cowork but Anthropic disputes risk

DGX agent

Security researchers at Armadin Inc. today detailed an attack chain that runs arbitrary commands as root inside the sandbox behind Anthropic PBC’s Claude Cowork, escaping the isolation layer, with a s

model-releasessiliconangle
1 Jul 2026
Model Releases

Artificial Intelligence in Sports: Insights from a Quantitative Survey among Sports Students in Germany about their Perceptions, Expectations, and Concerns regarding the Use of AI Tools

DGX agent

arXiv:2503.05785v2 Announce Type: replace-cross Abstract: Generative Artificial Intelligence (AI) tools such as ChatGPT, Copilot, or Gemini have a crucial impact on academic research and teaching. Emp

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

As generative AI tools continue to evolve, we believe it's more important than ever to know what's AI-generated and what isn't. That’s why @…

DGX agent

As generative AI tools continue to evolve, we believe it's more important than ever to know what's AI-generated and what isn't. That’s why @GoogleDeepMind launched SynthID in 2023—a technology that ad

model-releasesgoogle-ai--x
1 Jul 2026
Model Releases

AutoTrainess: Teaching Language Models to Improve Language Models Autonomously

DGX agent

arXiv:2606.31551v1 Announce Type: new Abstract: Training language models (LMs) remains a highly human-intensive process, even as frontier language model agents become increasingly capable at software

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

AxDafny: Agentic Verified Code Generation in Dafny

DGX agent

arXiv:2606.32007v1 Announce Type: new Abstract: We study agentic code generation in Dafny, where a model must generate both executable code and the proof artifacts for verification. We present AxDafny

model-releasesarxiv-cs-ai
1 Jul 2026
← Previous
1…141142143144145…472
Next →