AI Models
Track important developments across AI models, agents, research, business, hardware, robotics, regulation, safety and security — with source attribution and links to the original reporting.
Important AI developments right now.
Introducing Gemini 3.8 Live with Live Avatar
Read original ↗Gemini 3.8 text-to-speech says hello
Read original ↗Canopy: Exploiting Piecewise Smooth Tree Priors for Multi-Fidelity Bandits
arXiv:2609.30017v1 Announce Type: cross Abstract: Many LLM inference problems, including model routing, prefix-cache management, prompt trimming, and test-time search, can be viewed as optimization over a tree. This structure arises naturally from…
Read original ↗Google's "Call for Me" lets Gemini phone businesses for you
Google is testing "Call for Me," a feature that lets Gemini call businesses on a user's behalf. The article Google's "Call for Me" lets Gemini phone businesses for you appeared first on…
Canopy: Exploiting Piecewise Smooth Tree Priors for Multi-Fidelity Bandits
arXiv:2609.30017v1 Announce Type: cross Abstract: Many LLM inference problems, including model routing, prefix-cache management, prompt trimming, and test-time search, can be viewed as optimization over a tree. This structure arises naturally from…
When Search Becomes Memory: Accelerating Robot Design Discovery with Self-Evolving Skills
arXiv:2605.25832v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used as proposal generators for evolutionary robot design, yet most loops remain memoryless: simulator results shape the next population but are…
Pistis Technical Report
arXiv:2609.28554v1 Announce Type: new Abstract: We introduce the Pistis model family, comprising 27B- and 9B-parameter multimodal large language models built on Qwen3.6 and Qwen3.5, respectively, and developed through a general and scalable…
BaseCamp --- An Agentic AI Framework for Automating DNA Sequencing Data Pipelines
arXiv:2609.28557v1 Announce Type: new Abstract: DNA sequencing pipelines, spanning quality control, alignment, variant calling, and annotation, are now reliably executed by workflow management systems that orchestrate established bioinformatics tools at scale. What…
HERMES: A Holistic End-to-End Risk-Aware Multimodal Embodied System with Vision-Language Models for Long-Tail Autonomous Driving
arXiv:2602.00993v3 Announce Type: replace-cross Abstract: End-to-end autonomous driving models increasingly benefit from large vision-language models for semantic understanding, yet safe and reliable planning under long-tail conditions remains challenging, particularly in mixed-traffic environments…
Decoupling Knowledge and Privacy: Post-Task Self-Distillation Replay for LLM Continual Learning
arXiv:2609.29711v1 Announce Type: cross Abstract: Privacy-preserving continual learning (PPCL) must reduce the reproduction of sensitive content while retaining useful knowledge across sequential tasks. Formal privacy guarantees characterize randomized mechanisms, whereas operational output…
DEEPO: Dual-Entropy Enhanced Policy Optimization for Hallucination in MLLMs
arXiv:2609.28570v1 Announce Type: new Abstract: Reinforcement learning (RL) is widely used to sharpen reasoning in multimodal large language models (MLLMs), yet its effect on hallucination is uneven. We trace this to two…
Gemini 3.8 Live with Live Avatar gives Google’s AI a face
Google's new Gemini 3.8 Live update lets users have conversations with the model while watching an animated AI persona respond in real time. The "Live Avatar" will lip-sync and show different facial…
Introducing Gemini 3.8 Live with Live Avatar
Gemini 3.8 text-to-speech says hello
How invideo improves color grading 3x with GPT‑6 Astra
With GPT‑6 Astra, invideo plans edits with greater precision, improves color correction and grading threefold, and produces 50 custom effects in one day.
Parallel cut research time and cost in half with GPT‑6 Astra
GPT‑6 Astra allowed Parallel’s agents to research and synthesize labor-market data in half the time and at half the cost vs. prior models.
Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
CZITAPP shows headlines and concise feed summaries, attributes every source and links to the original publisher. Source type is descriptive, not a numerical truth score.