Gemini Interactions API

জেমিনি ইন্টারঅ্যাকশনস এপিআই ডেভেলপারদের জেমিনি মডেল ব্যবহার করে জেনারেটিভ এআই অ্যাপ্লিকেশন তৈরি করার সুযোগ দেয়। জেমিনি হলো আমাদের সবচেয়ে সক্ষম মডেল, যা একেবারে গোড়া থেকে মাল্টিমোডাল হওয়ার জন্য তৈরি করা হয়েছে। এটি ভাষা, ছবি, অডিও, ভিডিও এবং কোড সহ বিভিন্ন ধরণের তথ্যকে সাধারণীকরণ করতে, নির্বিঘ্নে বুঝতে, সেগুলোর মধ্যে কাজ করতে এবং একত্রিত করতে পারে। আপনি টেক্সট ও ছবির মধ্যে যুক্তিনির্মাণ, কন্টেন্ট তৈরি, ডায়ালগ এজেন্ট, সারসংক্ষেপ ও শ্রেণিবিন্যাস সিস্টেম এবং আরও অনেক কিছুর মতো ক্ষেত্রে জেমিনি এপিআই ব্যবহার করতে পারেন।

এপিআই সংস্করণ: v1beta v1

একটি মিথস্ক্রিয়া তৈরি করা

https://generativelanguage.googleapis.com/v1beta/interactions- এ পোস্ট করুন

একটি নতুন মিথস্ক্রিয়া তৈরি করে।

পাথ / কোয়েরি প্যারামিটার

api_version স্ট্রিং (আবশ্যক)

এপিআই-এর কোন সংস্করণটি ব্যবহার করতে হবে।

অনুরোধকারী শরীর

অনুরোধের মূল অংশে নিম্নলিখিত কাঠামোসহ ডেটা থাকে:

মডেল মডেলঅপশন (ঐচ্ছিক)

ইন্টারঅ্যাকশনটি তৈরি করতে ব্যবহৃত `মডেল`-এর নাম।
`agent` প্রদান করা না হলে এটি আবশ্যক।

যে মডেলটি আপনার প্রম্পটটি সম্পূর্ণ করবে।\n\nঅতিরিক্ত বিবরণের জন্য [মডেলসমূহ](https://ai.google.dev/gemini-api/docs/models) দেখুন।

সম্ভাব্য মান

  • models/gemini-2.5-flash-lite

    আমাদের সবচেয়ে ছোট এবং সবচেয়ে সাশ্রয়ী মডেল, যা ব্যাপক ব্যবহারের জন্য নির্মিত।

  • models/gemini-2.5-flash-image

    আমাদের নিজস্ব ইমেজ জেনারেশন মডেলটি গতি, নমনীয়তা এবং প্রাসঙ্গিকতা বোঝার জন্য অপ্টিমাইজ করা হয়েছে। টেক্সট ইনপুট এবং আউটপুটের মূল্য ২.৫ ফ্ল্যাশের সমান।

  • models/gemini-3.1-flash-lite

    আমাদের সবচেয়ে সাশ্রয়ী মডেল, যা বিপুল পরিমাণ এজেন্টিক কাজ, অনুবাদ এবং সাধারণ ডেটা প্রক্রিয়াকরণের জন্য অপ্টিমাইজ করা হয়েছে।

  • models/gemini-3.1-flash-image

    ফ্ল্যাশের গতির দক্ষতা এবং বাস্তবতার উপর ভিত্তি করে তৈরির ক্ষমতাসহ পেশাদার স্তরের ভিজ্যুয়াল ইন্টেলিজেন্স।

  • models/gemini-3.5-flash

    এজেন্টিক এবং কোডিং কার্যক্রমে ধারাবাহিক অগ্রণী পারফরম্যান্সের জন্য আমাদের সবচেয়ে বুদ্ধিমান মডেল।

  • models/gemini-3.6-flash

    এজেন্টিক এবং কোডিং কার্যক্রমে ধারাবাহিক অগ্রণী পারফরম্যান্সের জন্য আমাদের সবচেয়ে বুদ্ধিমান মডেল।

  • models/gemini-3.7-flash

    এজেন্টিক এবং কোডিং কার্যক্রমে ধারাবাহিক অগ্রণী পারফরম্যান্সের জন্য আমাদের সবচেয়ে বুদ্ধিমান মডেল।

এজেন্ট এজেন্টঅপশন (ঐচ্ছিক)

ইন্টারঅ্যাকশনটি তৈরি করতে ব্যবহৃত 'এজেন্ট'-এর নাম।
`model` প্রদান করা না হলে এটি আবশ্যক।

যে এজেন্টের সাথে যোগাযোগ করতে হবে।

সম্ভাব্য মান

  • deep-research-pro-preview-12-2025

    জেমিনি ডিপ রিসার্চ এজেন্ট

  • deep-research-preview-04-2026

    জেমিনি ডিপ রিসার্চ এজেন্ট

  • deep-research-max-preview-04-2026

    জেমিনি ডিপ রিসার্চ ম্যাক্স এজেন্ট

  • antigravity-preview-05-2026

    যুক্তিবোধ, ফাইল পরিচালনা এবং টুল ব্যবহারের প্রয়োজন হয় এমন একাধিক ধাপের কাজ সম্পাদন করতে অ্যান্টিগ্র্যাভিটি পরিচালিত এজেন্টটি ব্যবহার করুন।

ইনপুট কন্টেন্ট অথবা অ্যারে ( কন্টেন্ট ) অথবা অ্যারে ( ধাপ ) অথবা স্ট্রিং (আবশ্যক)

মিথস্ক্রিয়ার জন্য প্রয়োজনীয় উপাদানসমূহ (যা মডেল এবং এজেন্ট উভয়ের জন্যই প্রযোজ্য)।

সিস্টেম_নির্দেশনা স্ট্রিং (ঐচ্ছিক)

মিথস্ক্রিয়ার জন্য সিস্টেম নির্দেশাবলী।

সরঞ্জাম অ্যারে ( টুল ) (ঐচ্ছিক)

ইন্টারঅ্যাকশনের সময় মডেলটি যেসব টুল ডিক্লারেশন কল করতে পারে, তার একটি তালিকা।

response_format ResponseFormat অথবা অ্যারে ( ResponseFormat ) (ঐচ্ছিক)

এটি নিশ্চিত করে যে তৈরি হওয়া প্রতিক্রিয়াটি একটি JSON অবজেক্ট হবে যা এই ফিল্ডে নির্দিষ্ট করা JSON স্কিমা মেনে চলে।

স্ট্রিম বুলিয়ান (ঐচ্ছিক)

শুধুমাত্র ইনপুট। কথোপকথনটি স্ট্রিম করা হবে কিনা।

বুলিয়ান সংরক্ষণ করুন (ঐচ্ছিক)

শুধুমাত্র ইনপুট। প্রতিক্রিয়া এবং অনুরোধটি পরবর্তীতে পুনরুদ্ধারের জন্য সংরক্ষণ করা হবে কিনা।

পটভূমি বুলিয়ান (ঐচ্ছিক)

শুধুমাত্র ইনপুট। মডেল ইন্টারঅ্যাকশনটি ব্যাকগ্রাউন্ডে চালানো হবে কিনা।

generation_config GenerationConfig (ঐচ্ছিক)

মডেল কনফিগারেশন
মডেলের সাথে মিথস্ক্রিয়ার জন্য কনফিগারেশন প্যারামিটারসমূহ।
`agent_config`-এর বিকল্প। শুধুমাত্র তখনই প্রযোজ্য যখন `model` সেট করা থাকে।

মডেলের সাথে মিথস্ক্রিয়ার জন্য কনফিগারেশন প্যারামিটারসমূহ।

ক্ষেত্র

সর্বোচ্চ আউটপুট টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

প্রতিক্রিয়ায় অন্তর্ভুক্ত করার জন্য টোকেনের সর্বোচ্চ সংখ্যা।

বীজ পূর্ণসংখ্যা (ঐচ্ছিক)

পুনরুৎপাদনযোগ্যতার জন্য ডিকোডিং-এ ব্যবহৃত বীজ।

speech_config SpeakerConfig অথবা অ্যারে (SpeechConfig) (ঐচ্ছিক)

ঐচ্ছিক। বক্তৃতা এবং একাধিক স্পিকারের কনফিগারেশন।

স্টপ_সিকোয়েন্স অ্যারে (স্ট্রিং) (ঐচ্ছিক)

অক্ষর অনুক্রমের একটি তালিকা যা আউটপুট ইন্টারঅ্যাকশন বন্ধ করে দেবে।

চিন্তার স্তর ( ঐচ্ছিক)

মডেলটি যে পরিমাণ চিন্তার টোকেন তৈরি করবে।

সম্ভাব্য মান

  • minimal

    খুব কম বা একেবারেই চিন্তা না করা।

  • low

    চিন্তার নিম্ন স্তর।

  • medium

    মাঝারি চিন্তার স্তর।

  • high

    উচ্চ চিন্তার স্তর।

চিন্তার সারসংক্ষেপ (ঐচ্ছিক)

উত্তরে চিন্তার সারাংশ অন্তর্ভুক্ত করা হবে কিনা।

সম্ভাব্য মান

  • auto

    স্বয়ংক্রিয় চিন্তার সারাংশ।

  • none

    চিন্তামূলক সারসংক্ষেপ নয়।

tool_choice ToolChoiceConfig অথবা enum (স্ট্রিং) (ঐচ্ছিক)

টুল পছন্দের কনফিগারেশন।

সম্ভাব্য মানসমূহ:

  • auto

    স্বয়ংক্রিয় টুল নির্বাচন।

  • any

    যেকোনো সরঞ্জাম পছন্দ।

  • none

    সরঞ্জাম পছন্দের কোনো সুযোগ নেই।

  • validated

    সরঞ্জামটির নির্বাচন যথার্থ প্রমাণিত।

agent_config DynamicAgentConfig (ঐচ্ছিক)

এজেন্ট কনফিগারেশন
এজেন্টের জন্য কনফিগারেশন।
`generation_config`-এর বিকল্প। শুধুমাত্র তখনই প্রযোজ্য যখন `agent` সেট করা থাকে।

ডাইনামিক এজেন্টদের জন্য কনফিগারেশন।

ক্ষেত্র

অবজেক্ট টাইপ (ঐচ্ছিক)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "dynamic" এ সেট করা থাকে।

লেবেল অবজেক্ট (ঐচ্ছিক)

অনুরোধটির জন্য ব্যবহারকারী-সংজ্ঞায়িত মেটাডেটা সহ লেবেলগুলি।

max_total_tokens স্ট্রিং (ঐচ্ছিক)

এজেন্ট রানের জন্য সর্বোচ্চ মোট টোকেন।

পূর্ববর্তী_ইন্টারঅ্যাকশন_আইডি স্ট্রিং (ঐচ্ছিক)

পূর্ববর্তী যোগাযোগের আইডি, যদি থাকে।

safety_settings অ্যারে (নিরাপত্তা সেটিংস) (ঐচ্ছিক)

মিথস্ক্রিয়ার জন্য নিরাপত্তা সেটিংস।

প্রতিক্রিয়া

একটি ইন্টারঅ্যাকশন রিসোর্স ফেরত দেয়।

সাধারণ অনুরোধ

উদাহরণ প্রতিক্রিয়া

{
  "created": "2025-11-26T12:25:15Z",
  "id": "v1_ChdPU0F4YWFtNkFwS2kxZThQZ05lbXdROBIXT1NBeGFhbTZBcEtpMWU4UGdOZW13UTg",
  "model": "gemini-3.6-flash",
  "object": "interaction",
  "status": "completed",
  "steps": [
    {
      "type": "model_output",
      "content": [
        {
          "type": "text",
          "text": "Hello! I'm functioning perfectly and ready to assist you.\n\nHow are you doing today?"
        }
      ]
    }
  ],
  "updated": "2025-11-26T12:25:15Z",
  "usage": {
    "input_tokens_by_modality": [
      {
        "modality": "text",
        "tokens": 7
      }
    ],
    "total_cached_tokens": 0,
    "total_input_tokens": 7,
    "total_output_tokens": 20,
    "total_thought_tokens": 22,
    "total_tokens": 49,
    "total_tool_use_tokens": 0
  }
}

মাল্টি-টার্ন

উদাহরণ প্রতিক্রিয়া

{
  "created": "2025-11-26T12:22:47Z",
  "id": "v1_ChdPU0F4YWFtNkFwS2kxZThQZ05lbXdROBIXT1NBeGFhbTZBcEtpMWU4UGdOZW13UTg",
  "model": "gemini-3.6-flash",
  "object": "interaction",
  "status": "completed",
  "steps": [
    {
      "type": "model_output",
      "content": [
        {
          "type": "text",
          "text": "The capital of France is Paris."
        }
      ]
    }
  ],
  "updated": "2025-11-26T12:22:47Z",
  "usage": {
    "input_tokens_by_modality": [
      {
        "modality": "text",
        "tokens": 50
      }
    ],
    "total_cached_tokens": 0,
    "total_input_tokens": 50,
    "total_output_tokens": 10,
    "total_thought_tokens": 0,
    "total_tokens": 60,
    "total_tool_use_tokens": 0
  }
}

ইমেজ ইনপুট

উদাহরণ প্রতিক্রিয়া

{
  "created": "2025-11-26T12:22:47Z",
  "id": "v1_ChdPU0F4YWFtNkFwS2kxZThQZ05lbXdROBIXT1NBeGFhbTZBcEtpMWU4UGdOZW13UTg",
  "model": "gemini-3.6-flash",
  "object": "interaction",
  "status": "completed",
  "steps": [
    {
      "type": "model_output",
      "content": [
        {
          "type": "text",
          "text": "A white humanoid robot with glowing blue eyes stands holding a red skateboard."
        }
      ]
    }
  ],
  "updated": "2025-11-26T12:22:47Z",
  "usage": {
    "input_tokens_by_modality": [
      {
        "modality": "text",
        "tokens": 10
      },
      {
        "modality": "image",
        "tokens": 258
      }
    ],
    "total_cached_tokens": 0,
    "total_input_tokens": 268,
    "total_output_tokens": 20,
    "total_thought_tokens": 0,
    "total_tokens": 288,
    "total_tool_use_tokens": 0
  }
}

ফাংশন কলিং

উদাহরণ প্রতিক্রিয়া

{
  "created": "2025-11-26T12:22:47Z",
  "id": "v1_ChdPU0F4YWFtNkFwS2kxZThQZ05lbXdROBIXT1NBeGFhbTZBcEtpMWU4UGdOZW13UTg",
  "model": "gemini-3.6-flash",
  "object": "interaction",
  "status": "requires_action",
  "steps": [
    {
      "name": "get_weather",
      "type": "function_call",
      "arguments": {
        "location": "Boston, MA"
      },
      "id": "gth23981"
    }
  ],
  "updated": "2025-11-26T12:22:47Z",
  "usage": {
    "input_tokens_by_modality": [
      {
        "modality": "text",
        "tokens": 100
      }
    ],
    "total_cached_tokens": 0,
    "total_input_tokens": 100,
    "total_output_tokens": 25,
    "total_thought_tokens": 0,
    "total_tokens": 125,
    "total_tool_use_tokens": 50
  }
}

একটি ইন্টারঅ্যাকশন বাতিল করা

পোস্ট https://generativelanguage.googleapis.com/v1beta/interactions/{id}/cancel

আইডি দ্বারা একটি ইন্টারঅ্যাকশন বাতিল করে। এটি শুধুমাত্র চলমান ব্যাকগ্রাউন্ড ইন্টারঅ্যাকশনগুলোর ক্ষেত্রে প্রযোজ্য।

পাথ / কোয়েরি প্যারামিটার

api_version স্ট্রিং (আবশ্যক)

এপিআই-এর কোন সংস্করণটি ব্যবহার করতে হবে।

আইডি স্ট্রিং (আবশ্যক)

বাতিল করার জন্য ইন্টারঅ্যাকশনটির অনন্য শনাক্তকারী।

প্রতিক্রিয়া

একটি ইন্টারঅ্যাকশন রিসোর্স ফেরত দেয়।

মিথস্ক্রিয়া বাতিল করুন

উদাহরণ প্রতিক্রিয়া

{
  "created": "2025-11-26T12:25:15Z",
  "id": "v1_ChdPU0F4YWFtNkFwS2kxZThQZ05lbXdROBIXT1NBeGFhbTZBcEtpMWU4UGdOZW13UTg",
  "model": "gemini-3.6-flash",
  "object": "interaction",
  "status": "cancelled",
  "updated": "2025-11-26T12:25:15Z"
}

একটি মিথস্ক্রিয়া পুনরুদ্ধার করা

https://generativelanguage.googleapis.com/v1beta/interactions/{id} থেকে পান

`Interaction.id`-এর উপর ভিত্তি করে একটিমাত্র ইন্টারঅ্যাকশনের সম্পূর্ণ বিবরণ পুনরুদ্ধার করে।

পাথ / কোয়েরি প্যারামিটার

api_version স্ট্রিং (আবশ্যক)

এপিআই-এর কোন সংস্করণটি ব্যবহার করতে হবে।

আইডি স্ট্রিং (আবশ্যক)

পুনরুদ্ধার করার জন্য ইন্টারঅ্যাকশনটির অনন্য শনাক্তকারী।

last_event_id স্ট্রিং (ঐচ্ছিক)

ঐচ্ছিক। সেট করা থাকলে, ইভেন্ট আইডি দ্বারা চিহ্নিত ইভেন্টের পরের চাঙ্ক থেকে ইন্টারঅ্যাকশন স্ট্রিম পুনরায় শুরু হয়। এটি শুধুমাত্র তখনই ব্যবহার করা যাবে যখন `stream` সত্য হবে।

স্ট্রিম বুলিয়ান (ঐচ্ছিক)

true-তে সেট করা হলে, তৈরি হওয়া কন্টেন্ট পর্যায়ক্রমে স্ট্রিম করা হবে।

ডিফল্ট মান: False

প্রতিক্রিয়া

একটি ইন্টারঅ্যাকশন রিসোর্স ফেরত দেয়।

মিথস্ক্রিয়া করুন

উদাহরণ প্রতিক্রিয়া

{
  "created": "2025-11-26T12:25:15Z",
  "id": "v1_ChdPU0F4YWFtNkFwS2kxZThQZ05lbXdROBIXT1NBeGFhbTZBcEtpMWU4UGdOZW13UTg",
  "model": "gemini-3.6-flash",
  "object": "interaction",
  "status": "completed",
  "steps": [
    {
      "type": "model_output",
      "content": [
        {
          "type": "text",
          "text": "I'm doing great, thank you for asking! How can I help you today?"
        }
      ]
    }
  ],
  "updated": "2025-11-26T12:25:15Z"
}

একটি ইন্টারঅ্যাকশন মুছে ফেলা

https://generativelanguage.googleapis.com/v1beta/interactions/{id} মুছে ফেলুন

আইডি দ্বারা ইন্টারঅ্যাকশনটি মুছে দেয়।

পাথ / কোয়েরি প্যারামিটার

api_version স্ট্রিং (আবশ্যক)

এপিআই-এর কোন সংস্করণটি ব্যবহার করতে হবে।

আইডি স্ট্রিং (আবশ্যক)

মুছে ফেলার জন্য ইন্টারঅ্যাকশনটির অনন্য শনাক্তকারী।

প্রতিক্রিয়া

সফল হলে, প্রতিক্রিয়াটি খালি থাকে।

মুছে ফেলুন

সম্পদ

মিথস্ক্রিয়া

মিথস্ক্রিয়া সম্পদ।

ক্ষেত্র

এজেন্ট এজেন্টঅপশন (ঐচ্ছিক)

ইন্টারঅ্যাকশনটি তৈরি করতে ব্যবহৃত 'এজেন্ট'-এর নাম।

যে এজেন্টের সাথে যোগাযোগ করতে হবে।

সম্ভাব্য মান

  • deep-research-pro-preview-12-2025

    জেমিনি ডিপ রিসার্চ এজেন্ট

  • deep-research-preview-04-2026

    জেমিনি ডিপ রিসার্চ এজেন্ট

  • deep-research-max-preview-04-2026

    জেমিনি ডিপ রিসার্চ ম্যাক্স এজেন্ট

  • antigravity-preview-05-2026

    যুক্তিবোধ, ফাইল পরিচালনা এবং টুল ব্যবহারের প্রয়োজন হয় এমন একাধিক ধাপের কাজ সম্পাদন করতে অ্যান্টিগ্র্যাভিটি পরিচালিত এজেন্টটি ব্যবহার করুন।

agent_config DynamicAgentConfig (ঐচ্ছিক)

এজেন্ট ইন্টারঅ্যাকশনের জন্য কনফিগারেশন প্যারামিটারসমূহ।

ডাইনামিক এজেন্টদের জন্য কনফিগারেশন।

ক্ষেত্র

অবজেক্ট টাইপ (ঐচ্ছিক)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "dynamic" এ সেট করা থাকে।

তৈরি করা স্ট্রিং (ঐচ্ছিক)

শুধুমাত্র আউটপুট। যে সময়ে প্রতিক্রিয়াটি তৈরি করা হয়েছিল, সেই সময়টি ISO 8601 ফরম্যাটে (YYYY-MM-DDThh:mm:ssZ) উল্লেখ করতে হবে।

ত্রুটি অ্যারে (ত্রুটি) (ঐচ্ছিক)

শুধুমাত্র আউটপুট। ইন্টারঅ্যাকশনে ডায়াগনস্টিক ত্রুটি / প্ল্যাটফর্ম ভুল রেকর্ড করা হয়েছে।

একটি ইন্টারঅ্যাকশন থেকে প্রাপ্ত ত্রুটি বার্তা।

ক্ষেত্র

কোড স্ট্রিং (ঐচ্ছিক)

একটি URI যা ত্রুটির ধরণ শনাক্ত করে।

বার্তা স্ট্রিং (ঐচ্ছিক)

মানুষের পাঠযোগ্য একটি ত্রুটি বার্তা।

আইডি স্ট্রিং (ঐচ্ছিক)

আবশ্যক। শুধুমাত্র আউটপুট। মিথস্ক্রিয়া সম্পন্ন করার জন্য একটি অনন্য শনাক্তকারী।

ডিফল্ট হলো:

ইনপুট কন্টেন্ট বা অ্যারে ( কন্টেন্ট ) বা অ্যারে ( ধাপ ) বা স্ট্রিং (ঐচ্ছিক)

মিথস্ক্রিয়ার জন্য ইনপুট।

লেবেল অবজেক্ট (ঐচ্ছিক)

অনুরোধটির জন্য ব্যবহারকারী-সংজ্ঞায়িত মেটাডেটা সহ লেবেলগুলি।

max_total_tokens স্ট্রিং (ঐচ্ছিক)

এজেন্ট রানের জন্য সর্বোচ্চ মোট টোকেন।

মডেল মডেলঅপশন (ঐচ্ছিক)

ইন্টারঅ্যাকশনটি তৈরি করতে ব্যবহৃত `মডেল`-এর নাম।

যে মডেলটি আপনার প্রম্পটটি সম্পূর্ণ করবে।\n\nঅতিরিক্ত বিবরণের জন্য [মডেলসমূহ](https://ai.google.dev/gemini-api/docs/models) দেখুন।

সম্ভাব্য মান

  • models/gemini-2.5-flash-lite

    আমাদের সবচেয়ে ছোট এবং সবচেয়ে সাশ্রয়ী মডেল, যা ব্যাপক ব্যবহারের জন্য নির্মিত।

  • models/gemini-2.5-flash-image

    আমাদের নিজস্ব ইমেজ জেনারেশন মডেলটি গতি, নমনীয়তা এবং প্রাসঙ্গিকতা বোঝার জন্য অপ্টিমাইজ করা হয়েছে। টেক্সট ইনপুট এবং আউটপুটের মূল্য ২.৫ ফ্ল্যাশের সমান।

  • models/gemini-3.1-flash-lite

    আমাদের সবচেয়ে সাশ্রয়ী মডেল, যা বিপুল পরিমাণ এজেন্টিক কাজ, অনুবাদ এবং সাধারণ ডেটা প্রক্রিয়াকরণের জন্য অপ্টিমাইজ করা হয়েছে।

  • models/gemini-3.1-flash-image

    ফ্ল্যাশের গতির দক্ষতা এবং বাস্তবতার উপর ভিত্তি করে তৈরির ক্ষমতাসহ পেশাদার স্তরের ভিজ্যুয়াল ইন্টেলিজেন্স।

  • models/gemini-3.5-flash

    এজেন্টিক এবং কোডিং কার্যক্রমে ধারাবাহিক অগ্রণী পারফরম্যান্সের জন্য আমাদের সবচেয়ে বুদ্ধিমান মডেল।

  • models/gemini-3.6-flash

    এজেন্টিক এবং কোডিং কার্যক্রমে ধারাবাহিক অগ্রণী পারফরম্যান্সের জন্য আমাদের সবচেয়ে বুদ্ধিমান মডেল।

  • models/gemini-3.7-flash

    এজেন্টিক এবং কোডিং কার্যক্রমে ধারাবাহিক অগ্রণী পারফরম্যান্সের জন্য আমাদের সবচেয়ে বুদ্ধিমান মডেল।

পূর্ববর্তী_ইন্টারঅ্যাকশন_আইডি স্ট্রিং (ঐচ্ছিক)

পূর্ববর্তী যোগাযোগের আইডি, যদি থাকে।

response_format ResponseFormat অথবা অ্যারে ( ResponseFormat ) (ঐচ্ছিক)

এটি নিশ্চিত করে যে তৈরি হওয়া প্রতিক্রিয়াটি একটি JSON অবজেক্ট হবে যা এই ফিল্ডে নির্দিষ্ট করা JSON স্কিমা মেনে চলে।

safety_settings অ্যারে (নিরাপত্তা সেটিংস) (ঐচ্ছিক)

মিথস্ক্রিয়ার জন্য নিরাপত্তা সেটিংস।

স্ট্যাটাস এনাম (স্ট্রিং) (ঐচ্ছিক)

আবশ্যক। শুধুমাত্র আউটপুট। মিথস্ক্রিয়ার অবস্থা।

সম্ভাব্য মানসমূহ:

  • in_progress

    আলোচনাটি চলছে।

  • requires_action

    এই মিথস্ক্রিয়ার জন্য ব্যবহারকারীর পক্ষ থেকে কোনো পদক্ষেপ বা ইনপুট প্রয়োজন।

  • completed

    মিথস্ক্রিয়াটি সম্পন্ন হয়েছে।

  • failed

    মিথস্ক্রিয়াটি ব্যর্থ হয়েছে।

  • cancelled

    আলাপচারিতাটি বাতিল করা হয়েছিল।

  • incomplete

    ইন্টারঅ্যাকশনটি সম্পন্ন হয়েছে, কিন্তু এতে অসম্পূর্ণ ফলাফল রয়েছে (যেমন max_tokens-এ পৌঁছানো)।

ধাপ অ্যারে ( ধাপ ) (ঐচ্ছিক)

শুধুমাত্র আউটপুট। প্রতিক্রিয়ার অন্তর্ভুক্ত হলে, মিথস্ক্রিয়াটি গঠনকারী ধাপগুলো।

সিস্টেম_নির্দেশনা স্ট্রিং (ঐচ্ছিক)

মিথস্ক্রিয়ার জন্য সিস্টেম নির্দেশাবলী।

সরঞ্জাম অ্যারে ( টুল ) (ঐচ্ছিক)

ইন্টারঅ্যাকশনের সময় মডেলটি যেসব টুল ডিক্লারেশন কল করতে পারে, তার একটি তালিকা।

আপডেট করা স্ট্রিং (ঐচ্ছিক)

শুধুমাত্র আউটপুট। যে সময়ে প্রতিক্রিয়াটি সর্বশেষ আপডেট করা হয়েছিল, সেই সময়টি ISO 8601 ফরম্যাটে (YYYY-MM-DDThh:mm:ssZ)।

ব্যবহার (ঐচ্ছিক )

শুধুমাত্র আউটপুট। ইন্টারঅ্যাকশন অনুরোধের টোকেন ব্যবহারের পরিসংখ্যান।

ইন্টারঅ্যাকশন অনুরোধের টোকেন ব্যবহারের পরিসংখ্যান।

ক্ষেত্র

cached_tokens_by_modality অ্যারে (ModalityTokens) (ঐচ্ছিক)

পদ্ধতি অনুসারে ক্যাশড টোকেন ব্যবহারের বিশদ বিবরণ।

একটি একক প্রতিক্রিয়া পদ্ধতির জন্য টোকেন সংখ্যা।

ক্ষেত্র

পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)

টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।

সম্ভাব্য মান

  • text

    এটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।

  • image

    এটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।

  • audio

    এটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।

  • video

    এটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।

  • document

    এটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।

টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

মোডালিটির জন্য টোকেনের সংখ্যা।

গ্রাউন্ডিং_টুল_কাউন্ট অ্যারে (GroundingToolCount) (ঐচ্ছিক)

গ্রাউন্ডিং টুলের সংখ্যা।

গ্রাউন্ডিং টুলের সংখ্যা গুরুত্বপূর্ণ।

ক্ষেত্র

পূর্ণসংখ্যা গণনা করুন (ঐচ্ছিক)

গ্রাউন্ডিং টুলের সংখ্যা গুরুত্বপূর্ণ।

টাইপ এনাম (স্ট্রিং) (ঐচ্ছিক)

গণনার সাথে সংশ্লিষ্ট গ্রাউন্ডিং টুলের ধরণ।

সম্ভাব্য মানসমূহ:

  • google_search

    গুগল ওয়েব সার্চ ও ইমেজ সার্চের মাধ্যমে ভিত্তি স্থাপন, এবং এন্টারপ্রাইজের জন্য ওয়েব ভিত্তি স্থাপন।

  • google_maps

    গুগল ম্যাপসের সাহায্যে পরিচিতি।

input_tokens_by_modality অ্যারে (ModalityTokens) (ঐচ্ছিক)

পদ্ধতি অনুসারে ইনপুট টোকেন ব্যবহারের বিশদ বিবরণ।

একটি একক প্রতিক্রিয়া পদ্ধতির জন্য টোকেন সংখ্যা।

ক্ষেত্র

পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)

টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।

সম্ভাব্য মান

  • text

    এটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।

  • image

    এটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।

  • audio

    এটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।

  • video

    এটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।

  • document

    এটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।

টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

মোডালিটির জন্য টোকেনের সংখ্যা।

আউটপুট_টোকেন_বাই_মোডালিটি অ্যারে (মোডালিটিটোকেন) (ঐচ্ছিক)

পদ্ধতি অনুসারে আউটপুট টোকেন ব্যবহারের বিশদ বিবরণ।

একটি একক প্রতিক্রিয়া পদ্ধতির জন্য টোকেন সংখ্যা।

ক্ষেত্র

পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)

টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।

সম্ভাব্য মান

  • text

    এটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।

  • image

    এটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।

  • audio

    এটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।

  • video

    এটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।

  • document

    এটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।

টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

মোডালিটির জন্য টোকেনের সংখ্যা।

tool_use_tokens_by_modality অ্যারে (ModalityTokens) (ঐচ্ছিক)

পদ্ধতি অনুসারে টুল-ব্যবহার টোকেন ব্যবহারের একটি বিশদ বিবরণ।

একটি একক প্রতিক্রিয়া পদ্ধতির জন্য টোকেন সংখ্যা।

ক্ষেত্র

পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)

টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।

সম্ভাব্য মান

  • text

    এটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।

  • image

    এটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।

  • audio

    এটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।

  • video

    এটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।

  • document

    এটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।

টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

মোডালিটির জন্য টোকেনের সংখ্যা।

মোট ক্যাশ করা টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

প্রম্পটের ক্যাশ করা অংশে থাকা টোকেনের সংখ্যা (ক্যাশ করা বিষয়বস্তু)।

মোট ইনপুট টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

প্রম্পটে (প্রসঙ্গ) টোকেনের সংখ্যা।

মোট আউটপুট টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

তৈরি হওয়া সমস্ত প্রতিক্রিয়া জুড়ে টোকেনের মোট সংখ্যা।

মোট_চিন্তা_টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

চিন্তন মডেলগুলোর জন্য চিন্তার টোকেনের সংখ্যা।

মোট টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

ইন্টারঅ্যাকশন অনুরোধের জন্য মোট টোকেন সংখ্যা (প্রম্পট + প্রতিক্রিয়া + অন্যান্য অভ্যন্তরীণ টোকেন)।

total_tool_use_tokens পূর্ণসংখ্যা (ঐচ্ছিক)

টুল ব্যবহারের নির্দেশনায় উপস্থিত টোকেনের সংখ্যা।

উদাহরণ

উদাহরণ

{
  "created": "2025-12-04T15:01:45Z",
  "id": "v1_ChdXS0l4YWZXTk9xbk0xZThQczhEcmlROBIXV0tJeGFmV05PcW5NMWU4UHM4RHJpUTg",
  "model": "gemini-3.6-flash",
  "object": "interaction",
  "status": "completed",
  "steps": [
    {
      "type": "model_output",
      "content": [
        {
          "type": "text",
          "text": "Hello! I'm doing well, functioning as expected. Thank you for asking! How are you doing today?"
        }
      ]
    }
  ],
  "updated": "2025-12-04T15:01:45Z",
  "usage": {
    "input_tokens_by_modality": [
      {
        "modality": "text",
        "tokens": 7
      }
    ],
    "total_cached_tokens": 0,
    "total_input_tokens": 7,
    "total_output_tokens": 23,
    "total_thought_tokens": 49,
    "total_tokens": 79,
    "total_tool_use_tokens": 0
  }
}

ডেটা মডেল

বিষয়বস্তু

প্রতিক্রিয়ার বিষয়বস্তু।

সম্ভাব্য প্রকার

অডিও কন্টেন্ট

একটি অডিও কন্টেন্ট ব্লক।

চ্যানেল পূর্ণসংখ্যা (ঐচ্ছিক)

অডিও চ্যানেলের সংখ্যা।

ডেটা স্ট্রিং (ঐচ্ছিক)

অডিও বিষয়বস্তু।

mime_type enum (string) (ঐচ্ছিক)

অডিওটির মাইম টাইপ।

সম্ভাব্য মানসমূহ:

  • audio/wav

    WAV অডিও ফরম্যাট

  • audio/mp3

    MP3 অডিও ফরম্যাট

  • audio/aiff

    AIFF অডিও ফরম্যাট

  • audio/aac

    AAC অডিও ফরম্যাট

  • audio/ogg

    OGG অডিও ফরম্যাট

  • audio/flac

    FLAC অডিও ফরম্যাট

  • audio/mpeg

    MPEG অডিও ফরম্যাট

  • audio/m4a

    M4A অডিও ফরম্যাট

  • audio/l16

    L16 অডিও ফরম্যাট

  • audio/opus

    OPUS অডিও ফরম্যাট

  • audio/alaw

    ALAW অডিও ফরম্যাট

  • audio/mulaw

    MULAW অডিও ফরম্যাট

নমুনা_হার পূর্ণসংখ্যা (ঐচ্ছিক)

অডিওটির স্যাম্পল রেট।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "audio" তে সেট করা থাকে।

ইউআরআই স্ট্রিং (ঐচ্ছিক)

অডিওটির URI।

ডকুমেন্টের বিষয়বস্তু

একটি ডকুমেন্ট কন্টেন্ট ব্লক।

ডেটা স্ট্রিং (ঐচ্ছিক)

নথির বিষয়বস্তু।

mime_type enum (string) (ঐচ্ছিক)

ডকুমেন্টটির মাইম টাইপ।

সম্ভাব্য মানসমূহ:

  • application/pdf

    পিডিএফ ডকুমেন্ট ফরম্যাট

  • text/csv

    CSV ডকুমেন্ট ফরম্যাট

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "document" এ সেট করা থাকে।

ইউআরআই স্ট্রিং (ঐচ্ছিক)

ডকুমেন্টটির URI।

ছবির বিষয়বস্তু

একটি চিত্র বিষয়বস্তু ব্লক।

ডেটা স্ট্রিং (ঐচ্ছিক)

ছবির বিষয়বস্তু।

mime_type enum (string) (ঐচ্ছিক)

ছবিটির মাইম টাইপ।

সম্ভাব্য মানসমূহ:

  • image/png

    PNG ছবির ফরম্যাট

  • image/jpeg

    JPEG ছবির ফরম্যাট

  • image/webp

    ওয়েবপি ছবির ফরম্যাট

  • image/heic

    HEIC ছবির ফরম্যাট

  • image/heif

    HEIF ছবির ফরম্যাট

  • image/gif

    জিআইএফ ছবির ফরম্যাট

  • image/bmp

    বিএমপি ছবির ফরম্যাট

  • image/tiff

    টিআইএফএফ ইমেজ ফরম্যাট

রেজোলিউশন মিডিয়ারেজোলিউশন (ঐচ্ছিক)

গণমাধ্যমের সংকল্প।

সম্ভাব্য মান

  • low

    নিম্ন রেজোলিউশন।

  • medium

    মাঝারি রেজোলিউশন।

  • high

    উচ্চ রেজোলিউশন।

  • ultra_high

    অতি উচ্চ রেজোলিউশন।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "image" হিসেবে সেট করা থাকে।

ইউআরআই স্ট্রিং (ঐচ্ছিক)

ছবিটির URI।

টেক্সট কন্টেন্ট

একটি টেক্সট কন্টেন্ট ব্লক।

টীকা অ্যারে (টীকা) (ঐচ্ছিক)

মডেল-সৃষ্ট কন্টেন্টের জন্য উদ্ধৃতি তথ্য।

মডেল-সৃষ্ট কন্টেন্টের জন্য উদ্ধৃতি তথ্য।

সম্ভাব্য প্রকার

ফাইল উদ্ধৃতি

ফাইল উদ্ধৃতি টীকা।

কাস্টম_মেটাডেটা অবজেক্ট (ঐচ্ছিক)

ব্যবহারকারী সংগৃহীত কনটেক্সট সম্পর্কে মেটাডেটা প্রদান করেছেন।

ডকুমেন্ট_ইউআরআই স্ট্রিং (ঐচ্ছিক)

ফাইলটির URI।

শেষ_সূচক পূর্ণসংখ্যা (ঐচ্ছিক)

আরোপিত অংশের সমাপ্তি, স্বতন্ত্র।

ফাইলের নাম স্ট্রিং (ঐচ্ছিক)

ফাইলটির নাম।

মিডিয়া_আইডি স্ট্রিং (ঐচ্ছিক)

ছবির উদ্ধৃতির ক্ষেত্রে, প্রযোজ্য হলে মিডিয়া আইডি।

পৃষ্ঠা_সংখ্যা পূর্ণসংখ্যা (ঐচ্ছিক)

উদ্ধৃত নথির পৃষ্ঠা নম্বর, যদি প্রযোজ্য হয়।

উৎস স্ট্রিং (ঐচ্ছিক)

পাঠ্যের একটি অংশের উৎস উল্লেখ করা হয়েছে।

শুরুর সূচক পূর্ণসংখ্যা (ঐচ্ছিক)

প্রতিক্রিয়ার যে অংশটি এই উৎসের সাথে সম্পর্কিত, এটি তার শুরু। সূচকটি অংশের শুরু নির্দেশ করে, যা বাইটে পরিমাপ করা হয়।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "file_citation" এ সেট করা থাকে।

স্থান উদ্ধৃতি

স্থান উদ্ধৃতি টীকা।

শেষ_সূচক পূর্ণসংখ্যা (ঐচ্ছিক)

আরোপিত অংশের সমাপ্তি, স্বতন্ত্র।

নাম স্ট্রিং (ঐচ্ছিক)

স্থানটির নাম।

স্থান_আইডি স্ট্রিং (ঐচ্ছিক)

স্থানটির আইডি, `places/{place_id}` ফরম্যাটে।

review_snippets অ্যারে (রিভিউ স্নিপেট) (ঐচ্ছিক)

গুগল ম্যাপসে কোনো নির্দিষ্ট স্থানের বৈশিষ্ট্য সম্পর্কে উত্তর তৈরি করতে ব্যবহৃত পর্যালোচনার অংশবিশেষ।

গুগল ম্যাপসের কোনো নির্দিষ্ট স্থানের বৈশিষ্ট্য সম্পর্কে একটি প্রশ্নের উত্তর দেয় এমন ব্যবহারকারী পর্যালোচনার একটি অংশ এখানে তুলে ধরা হয়েছে।

ক্ষেত্র

review_id স্ট্রিং (ঐচ্ছিক)

রিভিউ স্নিপেটটির আইডি।

শিরোনাম স্ট্রিং (ঐচ্ছিক)

পর্যালোচনার শিরোনাম।

ইউআরএল স্ট্রিং (ঐচ্ছিক)

গুগল ম্যাপস-এ ব্যবহারকারীর পর্যালোচনার সাথে সম্পর্কিত একটি লিঙ্ক।

শুরুর সূচক পূর্ণসংখ্যা (ঐচ্ছিক)

প্রতিক্রিয়ার যে অংশটি এই উৎসের সাথে সম্পর্কিত, এটি তার শুরু। সূচকটি অংশের শুরু নির্দেশ করে, যা বাইটে পরিমাপ করা হয়।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "place_citation" এ সেট করা থাকে।

ইউআরএল স্ট্রিং (ঐচ্ছিক)

স্থানটির URI রেফারেন্স।

ইউআরএল উদ্ধৃতি

ইউআরএল উদ্ধৃতি টীকা।

শেষ_সূচক পূর্ণসংখ্যা (ঐচ্ছিক)

আরোপিত অংশের সমাপ্তি, স্বতন্ত্র।

শুরুর সূচক পূর্ণসংখ্যা (ঐচ্ছিক)

প্রতিক্রিয়ার যে অংশটি এই উৎসের সাথে সম্পর্কিত, এটি তার শুরু। সূচকটি অংশের শুরু নির্দেশ করে, যা বাইটে পরিমাপ করা হয়।

শিরোনাম স্ট্রিং (ঐচ্ছিক)

ইউআরএল-এর শিরোনাম।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "url_citation" এ সেট করা থাকে।

ইউআরএল স্ট্রিং (ঐচ্ছিক)

ইউআরএল।

টেক্সট স্ট্রিং (আবশ্যক)

প্রয়োজনীয়। পাঠ্য বিষয়বস্তু।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "text" হিসেবে সেট করা থাকে।

উদাহরণ

অডিও

{
  "type": "audio",
  "data": "BASE64_ENCODED_AUDIO",
  "mime_type": "audio/wav"
}

নথি

{
  "type": "document",
  "data": "BASE64_ENCODED_DOCUMENT",
  "mime_type": "application/pdf"
}

ছবি

{
  "type": "image",
  "data": "BASE64_ENCODED_IMAGE",
  "mime_type": "image/png"
}

পাঠ্য

{
  "type": "text",
  "text": "Hello, how are you?"
}

সরঞ্জাম

একটি সরঞ্জাম যা মডেল ব্যবহার করতে পারে।

সম্ভাব্য প্রকার

কোডএক্সিকিউশন

এমন একটি টুল যা মডেল কোড কার্যকর করার জন্য ব্যবহার করতে পারে।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "code_execution" এ সেট করা থাকে।

ফাইলসার্চ

একটি টুল যা মডেল ফাইল অনুসন্ধানের জন্য ব্যবহার করতে পারে।

ফাইল_সার্চ_স্টোর_নাম অ্যারে (স্ট্রিং) (ঐচ্ছিক)

ফাইল সার্চ স্টোরটি অনুসন্ধানের জন্য নামগুলো সংরক্ষণ করে।

মেটাডেটা_ফিল্টার স্ট্রিং (ঐচ্ছিক)

সিমান্টিক রিট্রিভাল ডকুমেন্ট এবং চাঙ্কগুলিতে প্রয়োগ করার জন্য মেটাডেটা ফিল্টার।

শীর্ষ_k পূর্ণসংখ্যা (ঐচ্ছিক)

পুনরুদ্ধার করার জন্য শব্দার্থিক পুনরুদ্ধার খণ্ডের সংখ্যা।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "file_search" এ সেট করা থাকে।

ফাংশন

একটি সরঞ্জাম যা মডেল ব্যবহার করতে পারে।

বর্ণনা স্ট্রিং (ঐচ্ছিক)

ফাংশনটির বর্ণনা।

নাম স্ট্রিং (ঐচ্ছিক)

ফাংশনটির নাম।

প্যারামিটার অবজেক্ট (ঐচ্ছিক)

ফাংশনের প্যারামিটারগুলোর জন্য JSON স্কিমা।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "function" এ সেট করা থাকে।

গুগলম্যাপস

একটি টুল যা মডেলটি গুগল ম্যাপস চালু করার জন্য ব্যবহার করতে পারে।

enable_widget বুলিয়ান (ঐচ্ছিক)

রেসপন্সের টুল কল রেজাল্টে উইজেট কনটেক্সট টোকেন ফেরত দেওয়া হবে কিনা।

অক্ষাংশ সংখ্যা (ঐচ্ছিক)

ব্যবহারকারীর অবস্থানের অক্ষাংশ।

দ্রাঘিমাংশ সংখ্যা (ঐচ্ছিক)

ব্যবহারকারীর অবস্থানের দ্রাঘিমাংশ।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "google_maps" এ সেট করা থাকে।

গুগল অনুসন্ধান

একটি টুল যা মডেল গুগলে অনুসন্ধান করার জন্য ব্যবহার করতে পারে।

অনুসন্ধানের ধরণ অ্যারে (এনাম (স্ট্রিং)) (ঐচ্ছিক)

সক্ষম করার জন্য অনুসন্ধানের ভিত্তি স্থাপনের প্রকারভেদ।

সম্ভাব্য মানসমূহ:

  • web_search

    এই ফিল্ডটি সেট করলে ওয়েব সার্চ চালু হয়। শুধুমাত্র টেক্সট ফলাফল দেখানো হয়।

  • image_search

    এই ফিল্ডটি সেট করলে ইমেজ সার্চ চালু হয়। ইমেজের বাইটগুলো ফেরত দেওয়া হয়।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "google_search" এ সেট করা থাকে।

ইউআরএলপ্রসঙ্গ

একটি টুল যা মডেল ইউআরএল কনটেক্সট সংগ্রহ করতে ব্যবহার করতে পারে।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "url_context" এ সেট করা থাকে।

উদাহরণ

কোডএক্সিকিউশন

ফাইলসার্চ

ফাংশন

গুগলম্যাপস

গুগল অনুসন্ধান

ইউআরএলপ্রসঙ্গ

ইন্টারঅ্যাকশনএসএসইইভেন্ট

সম্ভাব্য প্রকার

পলিমরফিক ডিসক্রিমিনেটর: event_type

ত্রুটি ইভেন্ট

ত্রুটি ( ঐচ্ছিক )

কোনো বিবরণ দেওয়া হয়নি।

একটি ইন্টারঅ্যাকশন থেকে প্রাপ্ত ত্রুটি বার্তা।

ক্ষেত্র

কোড স্ট্রিং (ঐচ্ছিক)

একটি URI যা ত্রুটির ধরণ শনাক্ত করে।

বার্তা স্ট্রিং (ঐচ্ছিক)

মানুষের পাঠযোগ্য একটি ত্রুটি বার্তা।

ইভেন্ট_আইডি স্ট্রিং (ঐচ্ছিক)

এই ইভেন্ট থেকে ইন্টারঅ্যাকশন স্ট্রিম পুনরায় শুরু করার জন্য ব্যবহৃত ইভেন্ট_আইডি টোকেন।

ইভেন্টের ধরণ অবজেক্ট (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "error" তে সেট করা থাকে।

মিথস্ক্রিয়া সম্পন্ন ইভেন্ট

ইভেন্ট_আইডি স্ট্রিং (ঐচ্ছিক)

এই ইভেন্ট থেকে ইন্টারঅ্যাকশন স্ট্রিম পুনরায় শুরু করার জন্য ব্যবহৃত ইভেন্ট_আইডি টোকেন।

ইভেন্টের ধরণ অবজেক্ট (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "interaction.completed" এ সেট করা থাকে।

মিথস্ক্রিয়া InteractionSseEventInteraction (প্রয়োজনীয়)

স্ট্রিমের শেষে আংশিকভাবে সম্পন্ন ইন্টারঅ্যাকশন রিসোর্স নির্গত হয়।

ইন্টারঅ্যাকশন লাইফসাইকেল SSE ইভেন্ট দ্বারা আংশিক ইন্টারঅ্যাকশন রিসোর্স নির্গত হয়। স্ট্রিমিং লাইফসাইকেল পেলোডগুলো এমন ফিল্ড বাদ দিতে পারে যা শুধুমাত্র সম্পূর্ণ নন-স্ট্রিমিং ইন্টারঅ্যাকশন রেসপন্সে পাওয়া যায়।

ক্ষেত্র

এজেন্ট স্ট্রিং (ঐচ্ছিক)

যে এজেন্টের সাথে যোগাযোগ করতে হবে।

তৈরি করা স্ট্রিং (ঐচ্ছিক)

শুধুমাত্র আউটপুট। যে সময়ে প্রতিক্রিয়াটি তৈরি করা হয়েছিল, সেই সময়টি ISO 8601 ফরম্যাটে।

আইডি স্ট্রিং (ঐচ্ছিক)

আবশ্যক। শুধুমাত্র আউটপুট। মিথস্ক্রিয়া সম্পন্ন করার জন্য একটি অনন্য শনাক্তকারী।

মডেল স্ট্রিং (ঐচ্ছিক)

যে মডেলটি আপনার নির্দেশটি সম্পূর্ণ করবে।

অবজেক্ট স্ট্রিং (ঐচ্ছিক)

শুধুমাত্র আউটপুট। রিসোর্স টাইপ।

স্ট্যাটাস এনাম (স্ট্রিং) (ঐচ্ছিক)

আবশ্যক। শুধুমাত্র আউটপুট। মিথস্ক্রিয়ার অবস্থা।

সম্ভাব্য মানসমূহ:

  • in_progress

    আলোচনাটি চলছে।

  • requires_action

    এই মিথস্ক্রিয়ার জন্য ব্যবহারকারীর পক্ষ থেকে কোনো পদক্ষেপ বা ইনপুট প্রয়োজন।

  • completed

    মিথস্ক্রিয়াটি সম্পন্ন হয়েছে।

  • failed

    মিথস্ক্রিয়াটি ব্যর্থ হয়েছে।

  • cancelled

    আলাপচারিতাটি বাতিল করা হয়েছিল।

  • incomplete

    ইন্টারঅ্যাকশনটি সম্পন্ন হয়েছে, কিন্তু এতে অসম্পূর্ণ ফলাফল রয়েছে (যেমন max_tokens-এ পৌঁছানো)।

ধাপ অ্যারে ( ধাপ ) (ঐচ্ছিক)

শুধুমাত্র আউটপুট। মিথস্ক্রিয়াটি গঠনকারী ধাপগুলো, যদি এই ইভেন্টে অন্তর্ভুক্ত থাকে।

আপডেট করা স্ট্রিং (ঐচ্ছিক)

শুধুমাত্র আউটপুট। যে সময়ে প্রতিক্রিয়াটি সর্বশেষ ISO 8601 ফরম্যাটে আপডেট করা হয়েছিল।

ব্যবহার (ঐচ্ছিক )

শুধুমাত্র আউটপুট। ইন্টারঅ্যাকশন অনুরোধের টোকেন ব্যবহারের পরিসংখ্যান।

ইন্টারঅ্যাকশন অনুরোধের টোকেন ব্যবহারের পরিসংখ্যান।

ক্ষেত্র

cached_tokens_by_modality অ্যারে (ModalityTokens) (ঐচ্ছিক)

পদ্ধতি অনুসারে ক্যাশড টোকেন ব্যবহারের বিশদ বিবরণ।

একটি একক প্রতিক্রিয়া পদ্ধতির জন্য টোকেন সংখ্যা।

ক্ষেত্র

পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)

টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।

সম্ভাব্য মান

  • text

    এটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।

  • image

    এটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।

  • audio

    এটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।

  • video

    এটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।

  • document

    এটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।

টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

মোডালিটির জন্য টোকেনের সংখ্যা।

গ্রাউন্ডিং_টুল_কাউন্ট অ্যারে (GroundingToolCount) (ঐচ্ছিক)

গ্রাউন্ডিং টুলের সংখ্যা।

গ্রাউন্ডিং টুলের সংখ্যা গুরুত্বপূর্ণ।

ক্ষেত্র

পূর্ণসংখ্যা গণনা করুন (ঐচ্ছিক)

গ্রাউন্ডিং টুলের সংখ্যা গুরুত্বপূর্ণ।

টাইপ এনাম (স্ট্রিং) (ঐচ্ছিক)

গণনার সাথে সংশ্লিষ্ট গ্রাউন্ডিং টুলের ধরণ।

সম্ভাব্য মানসমূহ:

  • google_search

    গুগল ওয়েব সার্চ ও ইমেজ সার্চের মাধ্যমে ভিত্তি স্থাপন, এবং এন্টারপ্রাইজের জন্য ওয়েব ভিত্তি স্থাপন।

  • google_maps

    গুগল ম্যাপসের সাহায্যে পরিচিতি।

input_tokens_by_modality অ্যারে (ModalityTokens) (ঐচ্ছিক)

পদ্ধতি অনুসারে ইনপুট টোকেন ব্যবহারের বিশদ বিবরণ।

একটি একক প্রতিক্রিয়া পদ্ধতির জন্য টোকেন সংখ্যা।

ক্ষেত্র

পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)

টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।

সম্ভাব্য মান

  • text

    এটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।

  • image

    এটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।

  • audio

    এটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।

  • video

    এটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।

  • document

    এটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।

টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

মোডালিটির জন্য টোকেনের সংখ্যা।

আউটপুট_টোকেন_বাই_মোডালিটি অ্যারে (মোডালিটিটোকেন) (ঐচ্ছিক)

পদ্ধতি অনুসারে আউটপুট টোকেন ব্যবহারের বিশদ বিবরণ।

একটি একক প্রতিক্রিয়া পদ্ধতির জন্য টোকেন সংখ্যা।

ক্ষেত্র

পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)

টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।

সম্ভাব্য মান

  • text

    এটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।

  • image

    এটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।

  • audio

    এটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।

  • video

    এটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।

  • document

    এটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।

টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

মোডালিটির জন্য টোকেনের সংখ্যা।

tool_use_tokens_by_modality অ্যারে (ModalityTokens) (ঐচ্ছিক)

পদ্ধতি অনুসারে টুল-ব্যবহার টোকেন ব্যবহারের একটি বিশদ বিবরণ।

একটি একক প্রতিক্রিয়া পদ্ধতির জন্য টোকেন সংখ্যা।

ক্ষেত্র

পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)

টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।

সম্ভাব্য মান

  • text

    এটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।

  • image

    এটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।

  • audio

    এটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।

  • video

    এটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।

  • document

    এটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।

টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

মোডালিটির জন্য টোকেনের সংখ্যা।

মোট ক্যাশ করা টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

প্রম্পটের ক্যাশ করা অংশে থাকা টোকেনের সংখ্যা (ক্যাশ করা বিষয়বস্তু)।

মোট ইনপুট টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

প্রম্পটে (প্রসঙ্গ) টোকেনের সংখ্যা।

মোট আউটপুট টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

তৈরি হওয়া সমস্ত প্রতিক্রিয়া জুড়ে টোকেনের মোট সংখ্যা।

মোট_চিন্তা_টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

চিন্তন মডেলগুলোর জন্য চিন্তার টোকেনের সংখ্যা।

মোট টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

ইন্টারঅ্যাকশন অনুরোধের জন্য মোট টোকেন সংখ্যা (প্রম্পট + প্রতিক্রিয়া + অন্যান্য অভ্যন্তরীণ টোকেন)।

total_tool_use_tokens পূর্ণসংখ্যা (ঐচ্ছিক)

টুল ব্যবহারের নির্দেশনায় উপস্থিত টোকেনের সংখ্যা।

ইন্টারঅ্যাকশন তৈরি ইভেন্ট

ইভেন্ট_আইডি স্ট্রিং (ঐচ্ছিক)

এই ইভেন্ট থেকে ইন্টারঅ্যাকশন স্ট্রিম পুনরায় শুরু করার জন্য ব্যবহৃত ইভেন্ট_আইডি টোকেন।

ইভেন্টের ধরণ অবজেক্ট (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "interaction.created" এ সেট করা থাকে।

মিথস্ক্রিয়া InteractionSseEventInteraction (প্রয়োজনীয়)

স্ট্রিমটি তৈরি করার সময় আংশিক ইন্টারঅ্যাকশন রিসোর্স নির্গত হয়।

ইন্টারঅ্যাকশন লাইফসাইকেল SSE ইভেন্ট দ্বারা আংশিক ইন্টারঅ্যাকশন রিসোর্স নির্গত হয়। স্ট্রিমিং লাইফসাইকেল পেলোডগুলো এমন ফিল্ড বাদ দিতে পারে যা শুধুমাত্র সম্পূর্ণ নন-স্ট্রিমিং ইন্টারঅ্যাকশন রেসপন্সে পাওয়া যায়।

ক্ষেত্র

এজেন্ট স্ট্রিং (ঐচ্ছিক)

যে এজেন্টের সাথে যোগাযোগ করতে হবে।

তৈরি করা স্ট্রিং (ঐচ্ছিক)

শুধুমাত্র আউটপুট। যে সময়ে প্রতিক্রিয়াটি তৈরি করা হয়েছিল, সেই সময়টি ISO 8601 ফরম্যাটে।

আইডি স্ট্রিং (ঐচ্ছিক)

আবশ্যক। শুধুমাত্র আউটপুট। মিথস্ক্রিয়া সম্পন্ন করার জন্য একটি অনন্য শনাক্তকারী।

মডেল স্ট্রিং (ঐচ্ছিক)

যে মডেলটি আপনার নির্দেশটি সম্পূর্ণ করবে।

অবজেক্ট স্ট্রিং (ঐচ্ছিক)

শুধুমাত্র আউটপুট। রিসোর্স টাইপ।

স্ট্যাটাস এনাম (স্ট্রিং) (ঐচ্ছিক)

আবশ্যক। শুধুমাত্র আউটপুট। মিথস্ক্রিয়ার অবস্থা।

সম্ভাব্য মানসমূহ:

  • in_progress

    আলোচনাটি চলছে।

  • requires_action

    এই মিথস্ক্রিয়ার জন্য ব্যবহারকারীর পক্ষ থেকে কোনো পদক্ষেপ বা ইনপুট প্রয়োজন।

  • completed

    মিথস্ক্রিয়াটি সম্পন্ন হয়েছে।

  • failed

    মিথস্ক্রিয়াটি ব্যর্থ হয়েছে।

  • cancelled

    আলাপচারিতাটি বাতিল করা হয়েছিল।

  • incomplete

    ইন্টারঅ্যাকশনটি সম্পন্ন হয়েছে, কিন্তু এতে অসম্পূর্ণ ফলাফল রয়েছে (যেমন max_tokens-এ পৌঁছানো)।

ধাপ অ্যারে ( ধাপ ) (ঐচ্ছিক)

শুধুমাত্র আউটপুট। মিথস্ক্রিয়াটি গঠনকারী ধাপগুলো, যদি এই ইভেন্টে অন্তর্ভুক্ত থাকে।

আপডেট করা স্ট্রিং (ঐচ্ছিক)

শুধুমাত্র আউটপুট। যে সময়ে প্রতিক্রিয়াটি সর্বশেষ ISO 8601 ফরম্যাটে আপডেট করা হয়েছিল।

ব্যবহার (ঐচ্ছিক )

শুধুমাত্র আউটপুট। ইন্টারঅ্যাকশন অনুরোধের টোকেন ব্যবহারের পরিসংখ্যান।

ইন্টারঅ্যাকশন অনুরোধের টোকেন ব্যবহারের পরিসংখ্যান।

ক্ষেত্র

cached_tokens_by_modality অ্যারে (ModalityTokens) (ঐচ্ছিক)

পদ্ধতি অনুসারে ক্যাশড টোকেন ব্যবহারের বিশদ বিবরণ।

একটি একক প্রতিক্রিয়া পদ্ধতির জন্য টোকেন সংখ্যা।

ক্ষেত্র

পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)

টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।

সম্ভাব্য মান

  • text

    এটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।

  • image

    এটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।

  • audio

    এটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।

  • video

    এটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।

  • document

    এটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।

টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

মোডালিটির জন্য টোকেনের সংখ্যা।

গ্রাউন্ডিং_টুল_কাউন্ট অ্যারে (GroundingToolCount) (ঐচ্ছিক)

গ্রাউন্ডিং টুলের সংখ্যা।

গ্রাউন্ডিং টুলের সংখ্যা গুরুত্বপূর্ণ।

ক্ষেত্র

পূর্ণসংখ্যা গণনা করুন (ঐচ্ছিক)

গ্রাউন্ডিং টুলের সংখ্যা গুরুত্বপূর্ণ।

টাইপ এনাম (স্ট্রিং) (ঐচ্ছিক)

গণনার সাথে সংশ্লিষ্ট গ্রাউন্ডিং টুলের ধরণ।

সম্ভাব্য মানসমূহ:

  • google_search

    গুগল ওয়েব সার্চ ও ইমেজ সার্চের মাধ্যমে ভিত্তি স্থাপন, এবং এন্টারপ্রাইজের জন্য ওয়েব ভিত্তি স্থাপন।

  • google_maps

    গুগল ম্যাপসের সাহায্যে পরিচিতি।

input_tokens_by_modality অ্যারে (ModalityTokens) (ঐচ্ছিক)

পদ্ধতি অনুসারে ইনপুট টোকেন ব্যবহারের বিশদ বিবরণ।

একটি একক প্রতিক্রিয়া পদ্ধতির জন্য টোকেন সংখ্যা।

ক্ষেত্র

পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)

টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।

সম্ভাব্য মান

  • text

    এটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।

  • image

    এটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।

  • audio

    এটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।

  • video

    এটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।

  • document

    এটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।

টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

মোডালিটির জন্য টোকেনের সংখ্যা।

আউটপুট_টোকেন_বাই_মোডালিটি অ্যারে (মোডালিটিটোকেন) (ঐচ্ছিক)

পদ্ধতি অনুসারে আউটপুট টোকেন ব্যবহারের বিশদ বিবরণ।

একটি একক প্রতিক্রিয়া পদ্ধতির জন্য টোকেন সংখ্যা।

ক্ষেত্র

পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)

টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।

সম্ভাব্য মান

  • text

    এটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।

  • image

    এটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।

  • audio

    এটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।

  • video

    এটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।

  • document

    এটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।

টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

মোডালিটির জন্য টোকেনের সংখ্যা।

tool_use_tokens_by_modality অ্যারে (ModalityTokens) (ঐচ্ছিক)

পদ্ধতি অনুসারে টুল-ব্যবহার টোকেন ব্যবহারের একটি বিশদ বিবরণ।

একটি একক প্রতিক্রিয়া পদ্ধতির জন্য টোকেন সংখ্যা।

ক্ষেত্র

পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)

টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।

সম্ভাব্য মান

  • text

    এটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।

  • image

    এটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।

  • audio

    এটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।

  • video

    এটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।

  • document

    এটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।

টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

মোডালিটির জন্য টোকেনের সংখ্যা।

মোট ক্যাশ করা টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

প্রম্পটের ক্যাশ করা অংশে থাকা টোকেনের সংখ্যা (ক্যাশ করা বিষয়বস্তু)।

মোট ইনপুট টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

প্রম্পটে (প্রসঙ্গ) টোকেনের সংখ্যা।

মোট আউটপুট টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

তৈরি হওয়া সমস্ত প্রতিক্রিয়া জুড়ে টোকেনের মোট সংখ্যা।

মোট_চিন্তা_টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

চিন্তন মডেলগুলোর জন্য চিন্তার টোকেনের সংখ্যা।

মোট টোকেন পূর্ণসংখ্যা (ঐচ্ছিক)

ইন্টারঅ্যাকশন অনুরোধের জন্য মোট টোকেন সংখ্যা (প্রম্পট + প্রতিক্রিয়া + অন্যান্য অভ্যন্তরীণ টোকেন)।

total_tool_use_tokens পূর্ণসংখ্যা (ঐচ্ছিক)

টুল ব্যবহারের নির্দেশনায় উপস্থিত টোকেনের সংখ্যা।

ইন্টারঅ্যাকশন স্ট্যাটাস আপডেট

ইভেন্ট_আইডি স্ট্রিং (ঐচ্ছিক)

এই ইভেন্ট থেকে ইন্টারঅ্যাকশন স্ট্রিম পুনরায় শুরু করার জন্য ব্যবহৃত ইভেন্ট_আইডি টোকেন।

ইভেন্টের ধরণ অবজেক্ট (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "interaction.status_update" এ সেট করা থাকে।

interaction_id স্ট্রিং (আবশ্যক)

কোনো বিবরণ দেওয়া হয়নি।

স্ট্যাটাস এনাম (স্ট্রিং) (আবশ্যক)

কোনো বিবরণ দেওয়া হয়নি।

সম্ভাব্য মানসমূহ:

  • in_progress

    আলোচনাটি চলছে।

  • requires_action

    এই মিথস্ক্রিয়ার জন্য ব্যবহারকারীর পক্ষ থেকে কোনো পদক্ষেপ বা ইনপুট প্রয়োজন।

  • completed

    মিথস্ক্রিয়াটি সম্পন্ন হয়েছে।

  • failed

    মিথস্ক্রিয়াটি ব্যর্থ হয়েছে।

  • cancelled

    আলাপচারিতাটি বাতিল করা হয়েছিল।

  • incomplete

    ইন্টারঅ্যাকশনটি সম্পন্ন হয়েছে, কিন্তু এতে অসম্পূর্ণ ফলাফল রয়েছে (যেমন max_tokens-এ পৌঁছানো)।

স্টেপডেল্টা

ডেল্টা স্টেপডেল্টাডেটা (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সম্ভাব্য প্রকার

আর্গুমেন্টসডেল্টা

আর্গুমেন্ট স্ট্রিং (ঐচ্ছিক)

কোনো বিবরণ দেওয়া হয়নি।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "arguments_delta" তে সেট করা থাকে।

অডিওডেল্টা

চ্যানেল পূর্ণসংখ্যা (ঐচ্ছিক)

অডিও চ্যানেলের সংখ্যা।

ডেটা স্ট্রিং (ঐচ্ছিক)

কোনো বিবরণ দেওয়া হয়নি।

mime_type enum (string) (ঐচ্ছিক)

কোনো বিবরণ দেওয়া হয়নি।

সম্ভাব্য মানসমূহ:

  • audio/wav

    WAV অডিও ফরম্যাট

  • audio/mp3

    MP3 অডিও ফরম্যাট

  • audio/aiff

    AIFF অডিও ফরম্যাট

  • audio/aac

    AAC অডিও ফরম্যাট

  • audio/ogg

    OGG অডিও ফরম্যাট

  • audio/flac

    FLAC অডিও ফরম্যাট

  • audio/mpeg

    MPEG অডিও ফরম্যাট

  • audio/m4a

    M4A অডিও ফরম্যাট

  • audio/l16

    L16 অডিও ফরম্যাট

  • audio/opus

    OPUS অডিও ফরম্যাট

  • audio/alaw

    ALAW অডিও ফরম্যাট

  • audio/mulaw

    MULAW অডিও ফরম্যাট

নমুনা_হার পূর্ণসংখ্যা (ঐচ্ছিক)

অডিওটির স্যাম্পল রেট।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "audio" তে সেট করা থাকে।

ইউআরআই স্ট্রিং (ঐচ্ছিক)

কোনো বিবরণ দেওয়া হয়নি।

কোডএক্সিকিউশনকলডেল্টা

আর্গুমেন্টস CodeExecutionCallArguments (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

কোড নির্বাহের জন্য আর্গুমেন্টগুলো প্রেরণ করতে হবে।

ক্ষেত্র

কোড স্ট্রিং (ঐচ্ছিক)

যে কোডটি কার্যকর করা হবে।

ভাষা এনাম (স্ট্রিং) (ঐচ্ছিক)

`কোড`-এর প্রোগ্রামিং ভাষা।

সম্ভাব্য মানসমূহ:

  • python

    পাইথন >= ৩.১০, সাথে numpy এবং simpy উপলব্ধ।

স্বাক্ষর স্ট্রিং (ঐচ্ছিক)

ব্যাকএন্ড যাচাইকরণের জন্য একটি স্বাক্ষর হ্যাশ।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "code_execution_call" এ সেট করা থাকে।

কোডএক্সিকিউশনরেজাল্টডেল্টা

is_error বুলিয়ান (ঐচ্ছিক)

কোনো বিবরণ দেওয়া হয়নি।

ফলাফল স্ট্রিং (আবশ্যক)

কোনো বিবরণ দেওয়া হয়নি।

স্বাক্ষর স্ট্রিং (ঐচ্ছিক)

ব্যাকএন্ড যাচাইকরণের জন্য একটি স্বাক্ষর হ্যাশ।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "code_execution_result" এ সেট করা থাকে।

ডকুমেন্টডেল্টা

ডেটা স্ট্রিং (ঐচ্ছিক)

কোনো বিবরণ দেওয়া হয়নি।

mime_type enum (string) (ঐচ্ছিক)

কোনো বিবরণ দেওয়া হয়নি।

সম্ভাব্য মানসমূহ:

  • application/pdf

    পিডিএফ ডকুমেন্ট ফরম্যাট

  • text/csv

    CSV ডকুমেন্ট ফরম্যাট

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

সর্বদা "document" এ সেট করা থাকে।

ইউআরআই স্ট্রিং (ঐচ্ছিক)

কোনো বিবরণ দেওয়া হয়নি।

ফাইলসার্চকলডেল্টা

স্বাক্ষর স্ট্রিং (ঐচ্ছিক)

ব্যাকএন্ড যাচাইকরণের জন্য একটি স্বাক্ষর হ্যাশ।

অবজেক্ট টাইপ (প্রয়োজনীয়)

কোনো বিবরণ দেওয়া হয়নি।

Always set to "file_search_call" .

FileSearchResultDelta

result array (FileSearchResult) (required)

No description provided.

The result of the File Search.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "file_search_result" .

FunctionResultDelta

call_id string (required)

Required. ID to match the ID from the function call block.

is_error boolean (optional)

No description provided.

name string (optional)

No description provided.

result array ( ImageContent or TextContent ) or object or string (required)

No description provided.

type object (required)

No description provided.

Always set to "function_result" .

GoogleMapsCallDelta

arguments GoogleMapsCallArguments (optional)

The arguments to pass to the Google Maps tool.

The arguments to pass to the Google Maps tool.

ক্ষেত্র

queries array (string) (optional)

The queries to be executed.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "google_maps_call" .

GoogleMapsResultDelta

result array (GoogleMapsResult) (optional)

The results of the Google Maps.

The result of the Google Maps.

ক্ষেত্র

places array (Places) (optional)

The places that were found.

ক্ষেত্র

name string (optional)

Title of the place.

place_id string (optional)

The ID of the place, in `places/{place_id}` format.

review_snippets array (ReviewSnippet) (optional)

Snippets of reviews that are used to generate answers about the features of a given place in Google Maps.

Encapsulates a snippet of a user review that answers a question about the features of a specific place in Google Maps.

ক্ষেত্র

review_id string (optional)

The ID of the review snippet.

title string (optional)

Title of the review.

url string (optional)

A link that corresponds to the user review on Google Maps.

url string (optional)

URI reference of the place.

widget_context_token string (optional)

Resource name of the Google Maps widget context token.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "google_maps_result" .

GoogleSearchCallDelta

arguments GoogleSearchCallArguments (required)

No description provided.

The arguments to pass to Google Search.

ক্ষেত্র

queries array (string) (optional)

Web search queries for the following-up web search.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "google_search_call" .

GoogleSearchResultDelta

is_error boolean (optional)

No description provided.

result array (GoogleSearchResult) (required)

No description provided.

The result of the Google Search.

ক্ষেত্র

search_suggestions string (optional)

Web content snippet that can be embedded in a web page or an app webview.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "google_search_result" .

ImageDelta

data string (optional)

No description provided.

mime_type enum (string) (optional)

No description provided.

Possible values:

  • image/png

    PNG image format

  • image/jpeg

    JPEG image format

  • image/webp

    WebP image format

  • image/heic

    HEIC image format

  • image/heif

    HEIF image format

  • image/gif

    GIF image format

  • image/bmp

    BMP image format

  • image/tiff

    TIFF image format

resolution MediaResolution (optional)

The resolution of the media.

Possible values

  • low

    Low resolution.

  • medium

    Medium resolution.

  • high

    High resolution.

  • ultra_high

    Ultra high resolution.

type object (required)

No description provided.

Always set to "image" .

uri string (optional)

No description provided.

TextAnnotationDelta

annotations array (Annotation) (optional)

Citation information for model-generated content.

Citation information for model-generated content.

Possible Types

FileCitation

A file citation annotation.

custom_metadata object (optional)

User provided metadata about the retrieved context.

document_uri string (optional)

The URI of the file.

end_index integer (optional)

End of the attributed segment, exclusive.

file_name string (optional)

The name of the file.

media_id string (optional)

Media ID in-case of image citations, if applicable.

page_number integer (optional)

Page number of the cited document, if applicable.

source string (optional)

Source attributed for a portion of the text.

start_index integer (optional)

Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.

type object (required)

No description provided.

Always set to "file_citation" .

PlaceCitation

A place citation annotation.

end_index integer (optional)

End of the attributed segment, exclusive.

name string (optional)

Title of the place.

place_id string (optional)

The ID of the place, in `places/{place_id}` format.

review_snippets array (ReviewSnippet) (optional)

Snippets of reviews that are used to generate answers about the features of a given place in Google Maps.

Encapsulates a snippet of a user review that answers a question about the features of a specific place in Google Maps.

ক্ষেত্র

review_id string (optional)

The ID of the review snippet.

title string (optional)

Title of the review.

url string (optional)

A link that corresponds to the user review on Google Maps.

start_index integer (optional)

Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.

type object (required)

No description provided.

Always set to "place_citation" .

url string (optional)

URI reference of the place.

UrlCitation

A URL citation annotation.

end_index integer (optional)

End of the attributed segment, exclusive.

start_index integer (optional)

Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.

title string (optional)

The title of the URL.

type object (required)

No description provided.

Always set to "url_citation" .

url string (optional)

The URL.

type object (required)

No description provided.

Always set to "text_annotation_delta" .

TextDelta

text string (required)

No description provided.

type object (required)

No description provided.

Always set to "text" .

ThoughtSignatureDelta

signature string (optional)

Signature to match the backend source to be part of the generation.

type object (required)

No description provided.

Always set to "thought_signature" .

ThoughtSummaryDelta

content Content (optional)

A new summary item to be added to the thought.

type object (required)

No description provided.

Always set to "thought_summary" .

UrlContextCallDelta

arguments UrlContextCallArguments (required)

No description provided.

The arguments to pass to the URL context.

ক্ষেত্র

urls array (string) (optional)

The URLs to fetch.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "url_context_call" .

UrlContextResultDelta

is_error boolean (optional)

No description provided.

result array (UrlContextResult) (required)

No description provided.

The result of the URL context.

ক্ষেত্র

status enum (string) (optional)

The status of the URL retrieval.

Possible values:

  • success

    Url retrieval is successful.

  • error

    Url retrieval is failed due to error.

  • paywall

    Url retrieval is failed because the content is behind paywall.

  • unsafe

    Url retrieval is failed because the content is unsafe.

url string (optional)

The URL that was fetched.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "url_context_result" .

event_id string (optional)

The event_id token to be used to resume the interaction stream, from this event.

event_type object (required)

No description provided.

Always set to "step.delta" .

index integer (required)

No description provided.

metadata StepDeltaMetadata (optional)

No description provided.

Optional metadata accompanying ANY streamed event.

ক্ষেত্র

total_usage Usage (optional)

Statistics on the interaction request's token usage.

Statistics on the interaction request's token usage.

ক্ষেত্র

cached_tokens_by_modality array (ModalityTokens) (optional)

A breakdown of cached token usage by modality.

The token count for a single response modality.

ক্ষেত্র

modality ResponseModality (optional)

The modality associated with the token count.

Possible values

  • text

    Indicates the model should return text.

  • image

    Indicates the model should return images.

  • audio

    Indicates the model should return audio.

  • video

    Indicates the model should return video.

  • document

    Indicates the model should return documents.

tokens integer (optional)

Number of tokens for the modality.

grounding_tool_count array (GroundingToolCount) (optional)

Grounding tool count.

The number of grounding tool counts.

ক্ষেত্র

count integer (optional)

The number of grounding tool counts.

type enum (string) (optional)

The grounding tool type associated with the count.

Possible values:

  • google_search

    Grounding with Google Web Search and Image Search, & Web Grounding for Enterprise.

  • google_maps

    Grounding with Google Maps.

input_tokens_by_modality array (ModalityTokens) (optional)

A breakdown of input token usage by modality.

The token count for a single response modality.

ক্ষেত্র

modality ResponseModality (optional)

The modality associated with the token count.

Possible values

  • text

    Indicates the model should return text.

  • image

    Indicates the model should return images.

  • audio

    Indicates the model should return audio.

  • video

    Indicates the model should return video.

  • document

    Indicates the model should return documents.

tokens integer (optional)

Number of tokens for the modality.

output_tokens_by_modality array (ModalityTokens) (optional)

A breakdown of output token usage by modality.

The token count for a single response modality.

ক্ষেত্র

modality ResponseModality (optional)

The modality associated with the token count.

Possible values

  • text

    Indicates the model should return text.

  • image

    Indicates the model should return images.

  • audio

    Indicates the model should return audio.

  • video

    Indicates the model should return video.

  • document

    Indicates the model should return documents.

tokens integer (optional)

Number of tokens for the modality.

tool_use_tokens_by_modality array (ModalityTokens) (optional)

A breakdown of tool-use token usage by modality.

The token count for a single response modality.

ক্ষেত্র

modality ResponseModality (optional)

The modality associated with the token count.

Possible values

  • text

    Indicates the model should return text.

  • image

    Indicates the model should return images.

  • audio

    Indicates the model should return audio.

  • video

    Indicates the model should return video.

  • document

    Indicates the model should return documents.

tokens integer (optional)

Number of tokens for the modality.

total_cached_tokens integer (optional)

Number of tokens in the cached part of the prompt (the cached content).

total_input_tokens integer (optional)

Number of tokens in the prompt (context).

total_output_tokens integer (optional)

Total number of tokens across all the generated responses.

total_thought_tokens integer (optional)

Number of tokens of thoughts for thinking models.

total_tokens integer (optional)

Total token count for the interaction request (prompt + responses + other internal tokens).

total_tool_use_tokens integer (optional)

Number of tokens present in tool-use prompt(s).

StepStart

event_id string (optional)

The event_id token to be used to resume the interaction stream, from this event.

event_type object (required)

No description provided.

Always set to "step.start" .

index integer (required)

No description provided.

step Step (required)

No description provided.

StepStop

event_id string (optional)

The event_id token to be used to resume the interaction stream, from this event.

event_type object (required)

No description provided.

Always set to "step.stop" .

index integer (required)

No description provided.

step_usage Usage (optional)

Model usage stats for this specific step.

Statistics on the interaction request's token usage.

ক্ষেত্র

cached_tokens_by_modality array (ModalityTokens) (optional)

A breakdown of cached token usage by modality.

The token count for a single response modality.

ক্ষেত্র

modality ResponseModality (optional)

The modality associated with the token count.

Possible values

  • text

    Indicates the model should return text.

  • image

    Indicates the model should return images.

  • audio

    Indicates the model should return audio.

  • video

    Indicates the model should return video.

  • document

    Indicates the model should return documents.

tokens integer (optional)

Number of tokens for the modality.

grounding_tool_count array (GroundingToolCount) (optional)

Grounding tool count.

The number of grounding tool counts.

ক্ষেত্র

count integer (optional)

The number of grounding tool counts.

type enum (string) (optional)

The grounding tool type associated with the count.

Possible values:

  • google_search

    Grounding with Google Web Search and Image Search, & Web Grounding for Enterprise.

  • google_maps

    Grounding with Google Maps.

input_tokens_by_modality array (ModalityTokens) (optional)

A breakdown of input token usage by modality.

The token count for a single response modality.

ক্ষেত্র

modality ResponseModality (optional)

The modality associated with the token count.

Possible values

  • text

    Indicates the model should return text.

  • image

    Indicates the model should return images.

  • audio

    Indicates the model should return audio.

  • video

    Indicates the model should return video.

  • document

    Indicates the model should return documents.

tokens integer (optional)

Number of tokens for the modality.

output_tokens_by_modality array (ModalityTokens) (optional)

A breakdown of output token usage by modality.

The token count for a single response modality.

ক্ষেত্র

modality ResponseModality (optional)

The modality associated with the token count.

Possible values

  • text

    Indicates the model should return text.

  • image

    Indicates the model should return images.

  • audio

    Indicates the model should return audio.

  • video

    Indicates the model should return video.

  • document

    Indicates the model should return documents.

tokens integer (optional)

Number of tokens for the modality.

tool_use_tokens_by_modality array (ModalityTokens) (optional)

A breakdown of tool-use token usage by modality.

The token count for a single response modality.

ক্ষেত্র

modality ResponseModality (optional)

The modality associated with the token count.

Possible values

  • text

    Indicates the model should return text.

  • image

    Indicates the model should return images.

  • audio

    Indicates the model should return audio.

  • video

    Indicates the model should return video.

  • document

    Indicates the model should return documents.

tokens integer (optional)

Number of tokens for the modality.

total_cached_tokens integer (optional)

Number of tokens in the cached part of the prompt (the cached content).

total_input_tokens integer (optional)

Number of tokens in the prompt (context).

total_output_tokens integer (optional)

Total number of tokens across all the generated responses.

total_thought_tokens integer (optional)

Number of tokens of thoughts for thinking models.

total_tokens integer (optional)

Total token count for the interaction request (prompt + responses + other internal tokens).

total_tool_use_tokens integer (optional)

Number of tokens present in tool-use prompt(s).

usage Usage (optional)

Cumulative model usage stats from the start of the session.

Statistics on the interaction request's token usage.

ক্ষেত্র

cached_tokens_by_modality array (ModalityTokens) (optional)

A breakdown of cached token usage by modality.

The token count for a single response modality.

ক্ষেত্র

modality ResponseModality (optional)

The modality associated with the token count.

Possible values

  • text

    Indicates the model should return text.

  • image

    Indicates the model should return images.

  • audio

    Indicates the model should return audio.

  • video

    Indicates the model should return video.

  • document

    Indicates the model should return documents.

tokens integer (optional)

Number of tokens for the modality.

grounding_tool_count array (GroundingToolCount) (optional)

Grounding tool count.

The number of grounding tool counts.

ক্ষেত্র

count integer (optional)

The number of grounding tool counts.

type enum (string) (optional)

The grounding tool type associated with the count.

Possible values:

  • google_search

    Grounding with Google Web Search and Image Search, & Web Grounding for Enterprise.

  • google_maps

    Grounding with Google Maps.

input_tokens_by_modality array (ModalityTokens) (optional)

A breakdown of input token usage by modality.

The token count for a single response modality.

ক্ষেত্র

modality ResponseModality (optional)

The modality associated with the token count.

Possible values

  • text

    Indicates the model should return text.

  • image

    Indicates the model should return images.

  • audio

    Indicates the model should return audio.

  • video

    Indicates the model should return video.

  • document

    Indicates the model should return documents.

tokens integer (optional)

Number of tokens for the modality.

output_tokens_by_modality array (ModalityTokens) (optional)

A breakdown of output token usage by modality.

The token count for a single response modality.

ক্ষেত্র

modality ResponseModality (optional)

The modality associated with the token count.

Possible values

  • text

    Indicates the model should return text.

  • image

    Indicates the model should return images.

  • audio

    Indicates the model should return audio.

  • video

    Indicates the model should return video.

  • document

    Indicates the model should return documents.

tokens integer (optional)

Number of tokens for the modality.

tool_use_tokens_by_modality array (ModalityTokens) (optional)

A breakdown of tool-use token usage by modality.

The token count for a single response modality.

ক্ষেত্র

modality ResponseModality (optional)

The modality associated with the token count.

Possible values

  • text

    Indicates the model should return text.

  • image

    Indicates the model should return images.

  • audio

    Indicates the model should return audio.

  • video

    Indicates the model should return video.

  • document

    Indicates the model should return documents.

tokens integer (optional)

Number of tokens for the modality.

total_cached_tokens integer (optional)

Number of tokens in the cached part of the prompt (the cached content).

total_input_tokens integer (optional)

Number of tokens in the prompt (context).

total_output_tokens integer (optional)

Total number of tokens across all the generated responses.

total_thought_tokens integer (optional)

Number of tokens of thoughts for thinking models.

total_tokens integer (optional)

Total token count for the interaction request (prompt + responses + other internal tokens).

total_tool_use_tokens integer (optional)

Number of tokens present in tool-use prompt(s).

উদাহরণ

Error Event

{
  "error": {
    "code": "not_found",
    "message": "Failed to get completed interaction: Result not found."
  },
  "event_type": "error"
}

Interaction Completed

{
  "event_id": "evt_123",
  "event_type": "interaction.completed",
  "interaction": {
    "created": "2025-12-04T15:01:45Z",
    "id": "v1_ChdXS0l4YWZXTk9xbk0xZThQczhEcmlROBIXV0tJeGFmV05PcW5NMWU4UHM4RHJpUTg",
    "model": "gemini-3.6-flash",
    "status": "completed",
    "updated": "2025-12-04T15:01:45Z"
  }
}

Interaction Completed

{
  "event_id": "evt_123",
  "event_type": "interaction.completed",
  "interaction": {
    "created": "2025-12-04T15:01:45Z",
    "id": "v1_ChdXS0l4YWZXTk9xbk0xZThQczhEcmlROBIXV0tJeGFmV05PcW5NMWU4UHM4RHJpUTg",
    "model": "gemini-3-flash-preview",
    "object": "interaction",
    "status": "completed",
    "updated": "2025-12-04T15:01:45Z"
  }
}

Interaction Created

{
  "event_id": "evt_123",
  "event_type": "interaction.created",
  "interaction": {
    "created": "2025-12-04T15:01:45Z",
    "id": "v1_ChdXS0l4YWZXTk9xbk0xZThQczhEcmlROBIXV0tJeGFmV05PcW5NMWU4UHM4RHJpUTg",
    "model": "gemini-3.6-flash",
    "status": "in_progress",
    "updated": "2025-12-04T15:01:45Z"
  }
}

Interaction Created

{
  "event_id": "evt_123",
  "event_type": "interaction.created",
  "interaction": {
    "id": "v1_ChdXS0l4YWZXTk9xbk0xZThQczhEcmlROBIXV0tJeGFmV05PcW5NMWU4UHM4RHJpUTg",
    "model": "gemini-3-flash-preview",
    "object": "interaction",
    "status": "in_progress"
  }
}

Interaction Status Update

{
  "event_type": "interaction.status_update",
  "interaction_id": "v1_ChdTMjQ0YWJ5TUF1TzcxZThQdjRpcnFRcxIXUzI0NGFieU1BdU83MWU4UHY0aXJxUXM",
  "status": "in_progress"
}

Step Delta

{
  "delta": {
    "type": "text",
    "text": "Hello"
  },
  "event_type": "step.delta",
  "index": 0
}

Step Start

{
  "event_type": "step.start",
  "index": 0,
  "step": {
    "type": "model_output"
  }
}

Step Stop

{
  "event_type": "step.stop",
  "index": 0
}

ResponseFormat

Possible Types

AudioResponseFormat

Configuration for audio output format.

bit_rate integer (optional)

Bit rate in bits per second (bps). Only applicable for compressed formats (MP3, Opus).

delivery enum (string) (optional)

The delivery mode for the audio output.

Possible values:

  • inline

    Audio data is returned inline in the response.

  • uri

    Audio data is returned as a URI.

mime_type enum (string) (optional)

The MIME type of the audio output.

Possible values:

  • audio/mp3

    MP3 audio format.

  • audio/ogg_opus

    OGG Opus audio format.

  • audio/l16

    Raw PCM (L16) audio format.

  • audio/wav

    WAV audio format.

  • audio/alaw

    A-law audio format.

  • audio/mulaw

    Mu-law audio format.

sample_rate integer (optional)

Sample rate in Hz.

type object (required)

No description provided.

Always set to "audio" .

ImageResponseFormat

Configuration for image output format.

aspect_ratio enum (string) (optional)

The aspect ratio for the image output.

Possible values:

  • 1:1

    ১:১ আকৃতির অনুপাত।

  • 2:3

    2:3 aspect ratio.

  • 3:2

    3:2 aspect ratio.

  • 3:4

    3:4 aspect ratio.

  • 4:3

    4:3 aspect ratio.

  • 4:5

    ৪:৫ আকৃতি অনুপাত।

  • 5:4

    5:4 aspect ratio.

  • 9:16

    9:16 aspect ratio.

  • 16:9

    16:9 aspect ratio.

  • 21:9

    21:9 aspect ratio.

  • 1:8

    1:8 aspect ratio.

  • 8:1

    8:1 aspect ratio.

  • 1:4

    1:4 aspect ratio.

  • 4:1

    4:1 aspect ratio.

delivery enum (string) (optional)

The delivery mode for the image output.

Possible values:

  • inline

    Image data is returned inline in the response.

  • uri

    Image data is returned as a URI.

image_size enum (string) (optional)

The size of the image output.

Possible values:

  • 512

    512px image size.

  • 1K

    1K image size.

  • 2K

    2K image size.

  • 4K

    4K image size.

mime_type enum (string) (optional)

The MIME type of the image output.

Possible values:

  • image/jpeg

    JPEG image format.

type object (required)

No description provided.

Always set to "image" .

TextResponseFormat

Configuration for text output format.

mime_type enum (string) (optional)

The MIME type of the text output.

Possible values:

  • application/json

    JSON output format.

  • text/plain

    Plain text output format.

schema object (optional)

The JSON schema that the output should conform to. Only applicable when mime_type is application/json.

type object (required)

No description provided.

Always set to "text" .

VideoResponseFormat

Configuration for video output format.

aspect_ratio enum (string) (optional)

The aspect ratio for the video output.

Possible values:

  • 16:9

    16:9 aspect ratio.

  • 9:16

    9:16 aspect ratio.

delivery enum (string) (optional)

The delivery mode for the video output.

Possible values:

  • inline

    Video data is returned inline in the response.

  • uri

    Video data is returned as a URI.

duration string (optional)

The duration for the video output.

resolution enum (string) (optional)

The video output resolution. Defaults to 720p.

Possible values:

  • 360p

    360p resolution.

  • 720p

    720p resolution.

  • 1080p

    1080p resolution.

  • 4k

    4K resolution.

type object (required)

No description provided.

Always set to "video" .

উদাহরণ

Audio Output

{
  "type": "audio",
  "sample_rate": 24000
}

Image Output

{
  "type": "image",
  "aspect_ratio": "16:9",
  "image_size": "1K",
  "mime_type": "image/jpeg"
}

Text Output (JSON Schema)

{
  "type": "text",
  "mime_type": "application/json",
  "schema": {
    "type": "object",
    "properties": {
      "ingredients": {
        "type": "array",
        "items": {
          "type": "string"
        }
      },
      "recipe_name": {
        "type": "string"
      }
    },
    "required": [
      "ingredients",
      "recipe_name"
    ]
  }
}

VideoResponseFormat

No examples available for this type.

ধাপ

A step in the interaction.

Possible Types

CodeExecutionCallStep

Code execution call step.

arguments CodeExecutionCallStepArguments (optional)

The arguments to pass to the code execution.

The arguments to pass to the code execution.

ক্ষেত্র

code string (optional)

The code to be executed.

language enum (string) (optional)

Programming language of the `code`.

Possible values:

  • python

    Python >= 3.10, with numpy and simpy available.

id string (required)

Required. A unique ID for this specific tool call.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "code_execution_call" .

CodeExecutionResultStep

Code execution result step.

call_id string (required)

Required. ID to match the ID from the function call block.

is_error boolean (optional)

Whether the code execution resulted in an error.

result string (optional)

The output of the code execution.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "code_execution_result" .

FileSearchCallStep

File Search call step.

id string (required)

Required. A unique ID for this specific tool call.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "file_search_call" .

FileSearchResultStep

File Search result step.

call_id string (required)

Required. ID to match the ID from the function call block.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "file_search_result" .

FunctionCallStep

A function tool call step.

arguments object (required)

Required. The arguments to pass to the function.

id string (required)

Required. A unique ID for this specific tool call.

name string (required)

Required. The name of the tool to call.

type object (required)

No description provided.

Always set to "function_call" .

FunctionResultStep

Result of a function tool call.

call_id string (required)

Required. ID to match the ID from the function call block.

is_error boolean (optional)

Whether the tool call resulted in an error.

name string (optional)

The name of the tool that was called.

result array ( ImageContent or TextContent ) or object or string (required)

Required. The result of the tool call.

type object (required)

No description provided.

Always set to "function_result" .

GoogleMapsCallStep

Google Maps call step.

arguments GoogleMapsCallStepArguments (optional)

The arguments to pass to the Google Maps tool.

The arguments to pass to the Google Maps tool.

ক্ষেত্র

queries array (string) (optional)

The queries to be executed.

id string (required)

Required. A unique ID for this specific tool call.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "google_maps_call" .

GoogleMapsResultStep

Google Maps result step.

call_id string (required)

Required. ID to match the ID from the function call block.

result array (GoogleMapsResultItem) (optional)

No description provided.

The result of the Google Maps.

ক্ষেত্র

places array (GoogleMapsResultPlaces) (optional)

No description provided.

ক্ষেত্র

name string (optional)

No description provided.

place_id string (optional)

No description provided.

review_snippets array (ReviewSnippet) (optional)

No description provided.

Encapsulates a snippet of a user review that answers a question about the features of a specific place in Google Maps.

ক্ষেত্র

review_id string (optional)

The ID of the review snippet.

title string (optional)

Title of the review.

url string (optional)

A link that corresponds to the user review on Google Maps.

url string (optional)

No description provided.

widget_context_token string (optional)

No description provided.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "google_maps_result" .

GoogleSearchCallStep

Google Search call step.

arguments GoogleSearchCallStepArguments (optional)

The arguments to pass to Google Search.

The arguments to pass to Google Search.

ক্ষেত্র

queries array (string) (optional)

Web search queries for the following-up web search.

id string (required)

Required. A unique ID for this specific tool call.

search_type enum (string) (optional)

The type of search grounding enabled.

Possible values:

  • web_search

    Setting this field enables web search. Only text results are returned.

  • image_search

    Setting this field enables image search. Image bytes are returned.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "google_search_call" .

GoogleSearchResultStep

Google Search result step.

call_id string (required)

Required. ID to match the ID from the function call block.

is_error boolean (optional)

Whether the Google Search resulted in an error.

result array (GoogleSearchResultItem) (optional)

The results of the Google Search.

The result of the Google Search.

ক্ষেত্র

search_suggestions string (optional)

Web content snippet that can be embedded in a web page or an app webview.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "google_search_result" .

ModelOutputStep

Output generated by the model.

content array ( Content ) (optional)

No description provided.

type object (required)

No description provided.

Always set to "model_output" .

ThoughtStep

A thought step.

signature string (optional)

A signature hash for backend validation.

summary array ( Content ) (optional)

A summary of the thought.

type object (required)

No description provided.

Always set to "thought" .

UrlContextCallStep

URL context call step.

arguments UrlContextCallArguments (optional)

The arguments to pass to the URL context.

The arguments to pass to the URL context.

ক্ষেত্র

urls array (string) (optional)

The URLs to fetch.

id string (required)

Required. A unique ID for this specific tool call.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "url_context_call" .

UrlContextResultStep

URL context result step.

call_id string (required)

Required. ID to match the ID from the function call block.

is_error boolean (optional)

Whether the URL context resulted in an error.

result array (UrlContextResult) (optional)

The results of the URL context.

The result of the URL context.

ক্ষেত্র

status enum (string) (optional)

The status of the URL retrieval.

Possible values:

  • success

    Url retrieval is successful.

  • error

    Url retrieval is failed due to error.

  • paywall

    Url retrieval is failed because the content is behind paywall.

  • unsafe

    Url retrieval is failed because the content is unsafe.

url string (optional)

The URL that was fetched.

signature string (optional)

A signature hash for backend validation.

type object (required)

No description provided.

Always set to "url_context_result" .

UserInputStep

Input provided by the user.

content array ( Content ) (optional)

No description provided.

type object (required)

No description provided.

Always set to "user_input" .

উদাহরণ

CodeExecutionCallStep

{
  "type": "code_execution_call",
  "arguments": {
    "code": "print(sum(range(1, 11)))"
  },
  "id": "code_call_71021"
}

CodeExecutionResultStep

{
  "type": "code_execution_result",
  "call_id": "code_call_71021",
  "result": "55\n"
}

FileSearchCallStep

{
  "type": "file_search_call",
  "id": "file_call_88192"
}

FileSearchResultStep

{
  "type": "file_search_result",
  "call_id": "file_call_88192"
}

FunctionCallStep

{
  "name": "get_weather",
  "type": "function_call",
  "arguments": {
    "location": "Boston, MA"
  },
  "id": "call_98231"
}

FunctionResultStep

{
  "name": "get_weather",
  "type": "function_result",
  "call_id": "call_98231",
  "result": [
    {
      "type": "text",
      "text": "{\"weather\":\"sunny\"}"
    }
  ]
}

GoogleMapsCallStep

{
  "type": "google_maps_call",
  "arguments": {
    "latitude": 37.7749,
    "longitude": -122.4194
  },
  "id": "maps_call_39201"
}

GoogleMapsResultStep

{
  "type": "google_maps_result",
  "call_id": "maps_call_39201",
  "result": [
    {
      "name": "Golden Gate Park",
      "place_id": "ChIJIQBpAG2ahYAR9R7bNdTLg8M",
      "rating": 4.8
    }
  ]
}

GoogleSearchCallStep

{
  "type": "google_search_call",
  "arguments": {
    "query": "Who won the men's 100m in Paris 2024?"
  },
  "id": "search_call_19201"
}

GoogleSearchResultStep

{
  "type": "google_search_result",
  "call_id": "search_call_19201",
  "result": [
    {
      "title": "Paris 2024 Olympics: Noah Lyles wins men's 100m gold",
      "url": "https://olympics.com/en/news/paris-2024-noah-lyles-wins-mens-100m-gold",
      "snippet": "American Noah Lyles won the Olympic men's 100m gold medal in a photo finish."
    }
  ]
}

ModelOutputStep

{
  "type": "model_output",
  "content": [
    {
      "type": "text",
      "text": "The capital of France is Paris."
    }
  ]
}

ThoughtStep

{
  "type": "thought",
  "signature": "thought_sig_abcd1234",
  "summary": [
    {
      "type": "text",
      "text": "The model is searching Google for the capital of France."
    }
  ]
}

UrlContextCallStep

{
  "type": "url_context_call",
  "arguments": {
    "urls": [
      "https://www.example.com"
    ]
  },
  "id": "url_call_10219"
}

UrlContextResultStep

{
  "type": "url_context_result",
  "call_id": "url_call_10219",
  "result": [
    {
      "title": "Example Domain",
      "url": "https://www.example.com",
      "snippet": "This domain is for use in illustrative examples in documents."
    }
  ]
}

UserInputStep

{
  "type": "user_input",
  "content": [
    {
      "type": "text",
      "text": "What is the capital of France?"
    }
  ]
}

ToolChoiceConfig

The tool choice configuration containing allowed tools.

ক্ষেত্র

allowed_tools AllowedTools (optional)

The allowed tools.

The configuration for allowed tools.

ক্ষেত্র

mode enum (string) (optional)

The mode of the tool choice.

Possible values:

  • auto

    Auto tool choice.

  • any

    Any tool choice.

  • none

    No tool choice.

  • validated

    Validated tool choice.

tools array (string) (optional)

The names of the allowed tools.

উদাহরণ

উদাহরণ

{
  "allowed_tools": {
    "mode": "any",
    "tools": [
      "my_tool"
    ]
  }
}

ImageContent

An image content block.

ক্ষেত্র

data string (optional)

The image content.

mime_type enum (string) (optional)

The mime type of the image.

Possible values:

  • image/png

    PNG image format

  • image/jpeg

    JPEG image format

  • image/webp

    WebP image format

  • image/heic

    HEIC image format

  • image/heif

    HEIF image format

  • image/gif

    GIF image format

  • image/bmp

    BMP image format

  • image/tiff

    TIFF image format

resolution MediaResolution (optional)

The resolution of the media.

Possible values

  • low

    Low resolution.

  • medium

    Medium resolution.

  • high

    High resolution.

  • ultra_high

    Ultra high resolution.

type object (optional)

No description provided.

Always set to "image" .

uri string (optional)

The URI of the image.

উদাহরণ

ছবি

{
  "type": "image",
  "data": "BASE64_ENCODED_IMAGE",
  "mime_type": "image/png"
}

TextContent

A text content block.

ক্ষেত্র

annotations array (Annotation) (optional)

Citation information for model-generated content.

Citation information for model-generated content.

Possible Types

FileCitation

A file citation annotation.

custom_metadata object (optional)

User provided metadata about the retrieved context.

document_uri string (optional)

The URI of the file.

end_index integer (optional)

End of the attributed segment, exclusive.

file_name string (optional)

The name of the file.

media_id string (optional)

Media ID in-case of image citations, if applicable.

page_number integer (optional)

Page number of the cited document, if applicable.

source string (optional)

Source attributed for a portion of the text.

start_index integer (optional)

Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.

type object (required)

No description provided.

Always set to "file_citation" .

PlaceCitation

A place citation annotation.

end_index integer (optional)

End of the attributed segment, exclusive.

name string (optional)

Title of the place.

place_id string (optional)

The ID of the place, in `places/{place_id}` format.

review_snippets array (ReviewSnippet) (optional)

Snippets of reviews that are used to generate answers about the features of a given place in Google Maps.

Encapsulates a snippet of a user review that answers a question about the features of a specific place in Google Maps.

ক্ষেত্র

review_id string (optional)

The ID of the review snippet.

title string (optional)

Title of the review.

url string (optional)

A link that corresponds to the user review on Google Maps.

start_index integer (optional)

Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.

type object (required)

No description provided.

Always set to "place_citation" .

url string (optional)

URI reference of the place.

UrlCitation

A URL citation annotation.

end_index integer (optional)

End of the attributed segment, exclusive.

start_index integer (optional)

Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.

title string (optional)

The title of the URL.

type object (required)

No description provided.

Always set to "url_citation" .

url string (optional)

The URL.

text string (optional)

Required. The text content.

type object (optional)

No description provided.

Always set to "text" .

উদাহরণ

পাঠ্য

{
  "type": "text",
  "text": "Hello, how are you?"
}