Skip to content

GitHub Actions workflow discussion_answering.yml fails on Q&A discussion creation with a Vertex AI 404 for gemini-3.5-flash #6104

Description

@ftnext

Run: https://github.com/google/adk-python/actions/runs/27451592905
Discussion: #6103

Error:

google.genai.errors.ClientError: 404 NOT_FOUND.
Publisher Model `projects/***/locations/***/publishers/google/models/gemini-3.5-flash`
was not found or your project does not have access to it.

The failing step is Run Answering Script in ADK Answering Agent for Discussions. The workflow sets GOOGLE_GENAI_USE_VERTEXAI=1, and the model is hardcoded in:

  • contributing/samples/adk_team/adk_answering_agent/agent.py
  • contributing/samples/adk_team/adk_answering_agent/gemini_assistant/agent.py

This looks similar to #5696, but that issue covered triage workflows using Gemini Developer API. This failure is in the discussion answering workflow using Vertex AI.

Possible cause: GOOGLE_CLOUD_LOCATION may point to a region where gemini-3.5-flash is unavailable, or the Actions GCP project/service account does not have access to that model. Current docs list gemini-3.5-flash availability for global, us, and eu.

Activity

  1. added
    documentation[Component] This issue is related to documentation, it will be transferred to adk-docs
    on Jun 13, 2026
  2. added theissue type on Jun 13, 2026
  3. sanketpatil06 commented on Jun 16, 2026

    @sanketpatil06

    Hi @ftnext , Thanks for flagging this model routing issue — the availability gap for the hardcoded gemini-3.5-flash under regional endpoints is clearly confirmed, and running the workflow without targeting a global endpoint immediately fires the 404 with no other setup needed.

    The approach in PR #6113 (making the discussion-answering model configurable rather than hardcoding it, while keeping the default configuration flexible) looks like exactly the right fix. This allows users to either opt into widely supported models that exist in standard regional endpoints, or supply their own custom model parameters.

    Could you also confirm the configuration override holds for your local runner environment — specifically that you can successfully specify an alternate model (or apply the "global" location override to keep using gemini-3.5-flash) — and verify that this setup successfully unblocks execution on your end?

  4. added
    request clarification[Status] The maintainer need clarification or more information from the author
    on Jun 16, 2026
  5. ftnext commented on Jun 26, 2026

    @ftnext
    ContributorAuthor

    I tested this locally with a Vertex AI-enabled GCP project.

    Confirmed:

    • The LLM_MODEL_NAME override is applied to the discussion answering agent.
    • With LLM_MODEL_NAME=gemini-3.5-flash and GOOGLE_CLOUD_LOCATION=global, the actual answering agent loads with root_agent.model == "gemini-3.5-flash".

    The PR's default model of gemini-2.5-flash may not be safe for this agent's current mixed tool configuration.

    400 INVALID_ARGUMENT: Multiple tools are supported only when they are all search tools.

  6. removed
    request clarification[Status] The maintainer need clarification or more information from the author
    on Jun 29, 2026
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

Labels

documentation[Component] This issue is related to documentation, it will be transferred to adk-docs

Type

Projects

No projects

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions