Pixelship agent on Replicate, Gemini 3.6 Flash batch on OpenRouter, and agent tools for builders
TL;DR
New agentic image tools and trending multimodal models arrived alongside expanded Gemini agent features and deployment options for production use.
What shipped
On 28 July 2026 new agent architectures and models appeared on Replicate and Hugging Face while Google expanded Gemini agent capabilities and OpenRouter added batch access. These releases focus on practical inference and workflow integration rather than broad announcements.
Replicate new models
Pixelship: Appmeloncreator launched Pixelship on Replicate as an agentic chain that links an LLM with an image model to produce higher-quality images. It supports eight runs on Cog 0.21.0 and runs through the HTTP API or web playground for immediate testing by builders.
Hugging Face trending
Three models gained traction on the Hugging Face Hub. They cover speech recognition and multimodal image-text tasks with ready download and fine-tuning paths.
- •VibeVoice-ASR-BitNet Microsoft released VibeVoice-ASR-BitNet, a speech recognition model trending on Hugging Face and built with ggml. Developers can download it for audio transcription workflows.
- •Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-GGUF LuffyTheFox published a Qwen-based image-text-to-text model that is trending on Hugging Face with the hermes library. It supports fine-tuning for multimodal projects.
- •Kimi-K3 Unsloth released Kimi-K3, an image-text-to-text model trending on Hugging Face and built with transformers. It offers direct inference and fine-tuning for production pipelines.
Vendor launches
Google highlighted Gemini 3.6 Flash agent use cases and added new managed agent features in the Gemini API. Additional updates covered edge hardware and startup programs.
- •Gemini agents for dairy operations Google showed how a Michigan dairy farmer uses Gemini 3.6 Flash agents to manage daily farm tasks, giving SMB owners a concrete example of agent-driven workflow automation.
- •NVIDIA Jetson platform NVIDIA promoted the Jetson platform for compact edge AI builds, allowing developers to run models locally on handheld devices instead of cloud endpoints.
- •Google AI tools for nonprofits Google introduced training resources that help nonprofits adopt its AI tools for operational tasks without prior technical setup.
- •Gemini API Managed Agents Google added hooks and 3.6 Flash support to Managed Agents in the Gemini API so developers can deploy reliable agents in production environments.
- •AI Mode in Search Google expanded AI Mode in Search with features that help users plan offline activities such as booking tickets, showing practical consumer-facing AI output.
- •AI Startup Support Program Google and KDDI launched a program to fund AI-native startups in Japan, providing capital and technical resources for early-stage builders.
Product Hunt picks
Seven new tools appeared on Product Hunt that focus on agent workspaces, file handling, and content workflows. Most target direct use by non-coders or small teams.
- •Lamoom Lamoom lets users run agent apps inside Claude or list their own agents for others to use.
- •Pinery Prose Pinery Prose acts as an AI co-author for books where every suggested edit appears as an approvable diff.
- •Tag Your Photos Tag Your Photos generates keywords for Apple Photos entirely on-device on a Mac.
- •SUB/WAVE SUB/WAVE provides self-hosted radio with an AI DJ and a single shared stream.
- •Growth Opt Playbook Growth Opt Playbook converts campaign data into recommended next marketing actions.
- •Liminal Liminal supplies a shared workspace that functions as a second brain for an individual, agent, and team.
Other
Several infrastructure and model access updates appeared from Databricks, Google, AWS, and Together AI. They address agent spend controls, high-QPS search, and dedicated inference routing.
- •Unity AI Gateway Budgets Databricks described how it uses Unity AI Gateway Budgets to track and limit spend on its internal coding agents.
- •OlmoEarth Platform AllenAI released the OlmoEarth Platform for continent-scale satellite inference with automatic failure recovery across distributed compute.
- •Gemini 3.6 Flash batch on OpenRouter Google added Gemini 3.6 Flash batch to OpenRouter with 1,049k context at $0.75 per million input tokens, enabling cost-effective large-batch runs.
- •AgentCore Gateway MCP support AWS updated AgentCore Gateway to support the MCP 2026-07-28 spec for standardized agent connections.
- •Databricks AI Search high QPS Databricks published guidance on moving AI Search from prototype to production at high query rates for retail and voice use cases.
- •Gemini 3.5 Flash batch on OpenRouter Google listed Gemini 3.5 Flash batch on OpenRouter at $0.75 per million input and $4.50 per million output for testing.
- •Together AI Dedicated Model Inference Together AI explained its three-part resource model for endpoints, deployments, and capacity-aware routing in dedicated inference.
- •Gemini 3.6 Flash batch routing OpenRouter enabled routing for Gemini 3.6 Flash batch so builders can add the model ID to existing API setups without new infrastructure.
Industry news
Spur funding round: Spur Intelligence raised $200 million from Insight Partners for its bot-detection system that identifies legitimate human traffic.
What this means for you
For Vibe Builders: You can now test Pixelship on Replicate to chain an LLM and image model for better outputs without writing code. Lamoom and Liminal on Product Hunt give ready agent workspaces while Gemini 3.6 Flash batch on OpenRouter lowers the cost of running longer agent sessions. Start with one of the new Hugging Face trending models or the Google agent examples to replace manual image or search steps in your current workflow.
For Non-techies: Google showed a dairy farmer using Gemini agents to handle daily tasks, and new nonprofit training resources make similar tools easier to adopt. Product Hunt entries such as Tag Your Photos and SUB/WAVE run locally or self-hosted so you can add AI features to photos or audio without subscriptions. OpenRouter batch pricing for Gemini models keeps costs predictable when you need occasional large jobs.
For Developers: Databricks shared spend controls for coding agents and high-QPS patterns for AI Search while AWS aligned AgentCore Gateway with the latest MCP spec. Together AI and OpenRouter added dedicated routing and batch endpoints for Gemini 3.6 Flash that you can benchmark against existing stacks. Evaluate the new Replicate and Hugging Face releases against your current inference setup before integrating them into production pipelines.
What to watch next
Watch for production reliability numbers on Gemini 3.6 Flash batch and any follow-up MCP spec updates from AWS. Track whether Databricks or Together AI release new cost or QPS benchmarks this week.
Harsh’s take
The day mixed consumer-facing agent demos with infrastructure tweaks, yet few releases included side-by-side benchmarks against existing tools. Builders risk adding yet another endpoint without clear reliability data. The practical move is to pick one new model or agent feature, run a fixed test set against your current stack, and drop anything that fails on cost or latency within the first day of testing.
by Harsh Desai
Sources
Replicate new models
Hugging Face trending
- •VibeVoice-ASR-BitNet by microsoft trends on HuggingFace
- •Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-GGUF by LuffyTheFox trends on HuggingFace
- •Kimi-K3 by unsloth trends on HuggingFace
Vendor launches
- •How Gemini Flash agents are helping a Michigan dairy farmer
- •Powerful Compute So Compact, It’s Clutch: Build AI in Your Hand With NVIDIA Jetson
- •3 new ways nonprofits can put AI to work
- •Gemini API Managed Agents: 3.6 Flash, hooks, and more
- •5 ways AI Mode in Search helps you enjoy the real world
- •Google and KDDI are ready to back Japanese startups.
Product Hunt picks
Other
- •How Databricks manages its own coding agent spend with Unity AI Gateway Budgets
- •Laboratoria’s Mariana Costa Empowers Women in Tech
- •The OlmoEarth Platform: Geospatial inference at planetary scale
- •Google: Gemini 3.6 Flash (batch) now available on OpenRouter (1,049k context, $0.75/M in, $3.75/M out)
- •How AgentCore Gateway supports the MCP 2026-07-28 spec
- •From prototype to production: High QPS for Databricks AI Search
- •Google: Gemini 3.5 Flash (batch) added on OpenRouter
- •InferenceConfiguring Dedicated Model InferenceThe three-part resource model behind Together AI Dedicated Model Inference: endpoints, deployments, configs, and how capacity-aware routing ties them together.
Industry news
More AI news
- Daily RoundupKimi K3 on AI Gateway, mage-flow on Replicate, and agent tools for builders
Vendors added model access, regional routing, and Slack hooks while new image and agent products appeared on Replicate and Product Hunt.
- Weekly DigestHermes Agent 80% latency cuts and 51 updates, OpenClaw Mac app, and durable export tools
Hermes Agent rolled out dozens of stability, speed, and integration fixes across three days while OpenClaw added a Mac app and remote server catalog.
- Daily RoundupMage-Flow-Edit-Turbo trends, Stable Audio Open ships, plus inbox cleanup tools
Microsoft and Owensong models hit Hugging Face trends while Replicate adds audio generation and PureBox.ai targets Gmail cleanup amid ongoing global model discussions.