জেমিনি ইন্টারঅ্যাকশনস এপিআই ডেভেলপারদের জেমিনি মডেল ব্যবহার করে জেনারেটিভ এআই অ্যাপ্লিকেশন তৈরি করার সুযোগ দেয়। জেমিনি হলো আমাদের সবচেয়ে সক্ষম মডেল, যা একেবারে গোড়া থেকে মাল্টিমোডাল হওয়ার জন্য তৈরি করা হয়েছে। এটি ভাষা, ছবি, অডিও, ভিডিও এবং কোড সহ বিভিন্ন ধরণের তথ্যকে সাধারণীকরণ করতে, নির্বিঘ্নে বুঝতে, সেগুলোর মধ্যে কাজ করতে এবং একত্রিত করতে পারে। আপনি টেক্সট ও ছবির মধ্যে যুক্তিনির্মাণ, কন্টেন্ট তৈরি, ডায়ালগ এজেন্ট, সারসংক্ষেপ ও শ্রেণিবিন্যাস সিস্টেম এবং আরও অনেক কিছুর মতো ক্ষেত্রে জেমিনি এপিআই ব্যবহার করতে পারেন।
একটি মিথস্ক্রিয়া তৈরি করা
একটি নতুন মিথস্ক্রিয়া তৈরি করে।
পাথ / কোয়েরি প্যারামিটার
এপিআই-এর কোন সংস্করণটি ব্যবহার করতে হবে।
অনুরোধকারী শরীর
অনুরোধের মূল অংশে নিম্নলিখিত কাঠামোসহ ডেটা থাকে:
মডেল মডেলঅপশন (ঐচ্ছিক)
ইন্টারঅ্যাকশনটি তৈরি করতে ব্যবহৃত `মডেল`-এর নাম।
`agent` প্রদান করা না হলে এটি আবশ্যক।
সম্ভাব্য মান
-
gemini-2.5-flashআমাদের প্রথম হাইব্রিড রিজনিং মডেল যা ১ মিলিয়ন টোকেন কনটেক্সট উইন্ডো এবং থিংকিং বাজেট সমর্থন করে।
-
gemini-2.5-proআমাদের সর্বাধুনিক বহুমুখী মডেল, যা কোডিং এবং জটিল যুক্তিনির্ভর কাজে অত্যন্ত পারদর্শী।
-
gemma-4-26b-a4b-itজেমা 4 26B A4B IT
-
gemma-4-31b-itজেমা ৪ ৩১বি আইটি
-
gemini-flash-latestজেমিনি ফ্ল্যাশের সর্বশেষ সংস্করণ
-
gemini-flash-lite-latestজেমিনি ফ্ল্যাশ-লাইটের সর্বশেষ সংস্করণ
-
gemini-pro-latestজেমিনি প্রো-এর সর্বশেষ সংস্করণ
-
gemini-2.5-flash-liteআমাদের সবচেয়ে ছোট এবং সবচেয়ে সাশ্রয়ী মডেল, যা ব্যাপক ব্যবহারের জন্য নির্মিত।
-
gemini-2.5-flash-imageআমাদের নিজস্ব ইমেজ জেনারেশন মডেলটি গতি, নমনীয়তা এবং প্রাসঙ্গিকতা বোঝার জন্য অপ্টিমাইজ করা হয়েছে। টেক্সট ইনপুট এবং আউটপুটের মূল্য ২.৫ ফ্ল্যাশের সমান।
-
gemini-3-flash-previewগতির জন্য নির্মিত আমাদের সবচেয়ে বুদ্ধিমান মডেল, যা অত্যাধুনিক বুদ্ধিমত্তার সাথে উন্নত অনুসন্ধান এবং ভূমিতে স্থিতিশীলতার সমন্বয় ঘটায়।
-
gemini-3.1-pro-previewআমাদের সর্বাধুনিক অত্যাধুনিক রিজনিং মডেল, যা অভূতপূর্ব গভীরতা ও সূক্ষ্মতা এবং শক্তিশালী মাল্টিমোডাল আন্ডারস্ট্যান্ডিং ও কোডিং সক্ষমতাসম্পন্ন।
-
gemini-3.1-pro-preview-customtoolsকাস্টম টুল ব্যবহারের জন্য অপ্টিমাইজ করা জেমিনি ৩.১ প্রো প্রিভিউ।
-
gemini-3.1-flash-liteআমাদের সবচেয়ে সাশ্রয়ী মডেল, যা বিপুল পরিমাণ এজেন্টিক কাজ, অনুবাদ এবং সাধারণ ডেটা প্রক্রিয়াকরণের জন্য অপ্টিমাইজ করা হয়েছে।
-
gemini-3-pro-imageজেমিনি ৩ প্রো ইমেজ
-
nano-banana-pro-previewজেমিনি ৩ প্রো ছবির প্রিভিউ
-
gemini-3.1-flash-imageজেমিনি ৩.১ ফ্ল্যাশ চিত্র।
-
gemini-3.5-flashএজেন্টিক এবং কোডিং কার্যক্রমে ধারাবাহিক অগ্রণী পারফরম্যান্সের জন্য আমাদের সবচেয়ে বুদ্ধিমান মডেল।
-
gemini-3.6-flashএজেন্টিক এবং কোডিং কার্যক্রমে ধারাবাহিক অগ্রণী পারফরম্যান্সের জন্য আমাদের সবচেয়ে বুদ্ধিমান মডেল।
-
gemini-3.7-flashএজেন্টিক এবং কোডিং কার্যক্রমে ধারাবাহিক অগ্রণী পারফরম্যান্সের জন্য আমাদের সবচেয়ে বুদ্ধিমান মডেল।
-
lyria-3-clip-previewআমাদের স্বল্প-বিলম্বের সঙ্গীত তৈরির মডেলটি উচ্চ-মানের অডিও ক্লিপ এবং সুনির্দিষ্ট ছন্দ নিয়ন্ত্রণের জন্য অপ্টিমাইজ করা হয়েছে।
-
lyria-3-pro-previewগভীর সুরসৃষ্টিগত বোধসম্পন্ন আমাদের উন্নত, পূর্ণাঙ্গ গান তৈরির জেনারেটিভ মডেলটি, বিভিন্ন সঙ্গীত শৈলীতে সুনির্দিষ্ট কাঠামোগত নিয়ন্ত্রণ এবং জটিল রূপান্তরের জন্য সর্বোত্তমভাবে প্রস্তুত করা হয়েছে।
-
gemini-robotics-er-1.6-previewজেমিনি রোবোটিক্স-ইআর ১.৬ প্রিভিউ
-
gemini-robotics-er-2-previewজেমিনি রোবোটিক্স এমবডিড রিজনিং ২ প্রিভিউ
এজেন্ট এজেন্টঅপশন (ঐচ্ছিক)
ইন্টারঅ্যাকশনটি তৈরি করতে ব্যবহৃত 'এজেন্ট'-এর নাম।
`model` প্রদান করা না হলে এটি আবশ্যক।
সম্ভাব্য মান
-
deep-research-pro-preview-12-2025জেমিনি ডিপ রিসার্চ এজেন্ট
-
deep-research-preview-04-2026জেমিনি ডিপ রিসার্চ এজেন্ট
-
deep-research-max-preview-04-2026জেমিনি ডিপ রিসার্চ ম্যাক্স এজেন্ট
-
antigravity-preview-05-2026যুক্তিবোধ, ফাইল পরিচালনা এবং টুল ব্যবহারের প্রয়োজন হয় এমন একাধিক ধাপের কাজ সম্পাদন করতে অ্যান্টিগ্র্যাভিটি পরিচালিত এজেন্টটি ব্যবহার করুন।
মিথস্ক্রিয়ার জন্য প্রয়োজনীয় উপাদানসমূহ (যা মডেল এবং এজেন্ট উভয়ের জন্যই প্রযোজ্য)।
মিথস্ক্রিয়ার জন্য সিস্টেম নির্দেশাবলী।
ইন্টারঅ্যাকশনের সময় মডেলটি যেসব টুল ডিক্লারেশন কল করতে পারে, তার একটি তালিকা।
এটি নিশ্চিত করে যে তৈরি হওয়া প্রতিক্রিয়াটি একটি JSON অবজেক্ট হবে যা এই ফিল্ডে নির্দিষ্ট করা JSON স্কিমা মেনে চলে।
শুধুমাত্র ইনপুট। কথোপকথনটি স্ট্রিম করা হবে কিনা।
শুধুমাত্র ইনপুট। প্রতিক্রিয়া এবং অনুরোধটি পরবর্তীতে পুনরুদ্ধারের জন্য সংরক্ষণ করা হবে কিনা।
শুধুমাত্র ইনপুট। মডেল ইন্টারঅ্যাকশনটি ব্যাকগ্রাউন্ডে চালানো হবে কিনা।
generation_config GenerationConfig (ঐচ্ছিক)
মডেল কনফিগারেশন
মডেলের সাথে মিথস্ক্রিয়ার জন্য কনফিগারেশন প্যারামিটারসমূহ।
`agent_config`-এর বিকল্প। শুধুমাত্র তখনই প্রযোজ্য যখন `model` সেট করা থাকে।
ক্ষেত্র
প্রতিক্রিয়ায় অন্তর্ভুক্ত করার জন্য টোকেনের সর্বোচ্চ সংখ্যা।
পুনরুৎপাদনযোগ্যতার জন্য ডিকোডিং-এ ব্যবহৃত বীজ।
speech_config SpeakerConfig অথবা অ্যারে (SpeechConfig) (ঐচ্ছিক)
ঐচ্ছিক। বক্তৃতা এবং একাধিক স্পিকারের কনফিগারেশন।
ক্ষেত্র
স্পিকার অ্যারে (স্পিচকনফিগ) (ঐচ্ছিক)
স্বতন্ত্র স্পিকার কনফিগারেশন।
ক্ষেত্র
বক্তৃতার ভাষা।
বক্তার নাম অবশ্যই প্রম্পটে দেওয়া বক্তার নামের সাথে মিলতে হবে।
বক্তার কণ্ঠস্বর।
অক্ষর অনুক্রমের একটি তালিকা যা আউটপুট ইন্টারঅ্যাকশন বন্ধ করে দেবে।
চিন্তার স্তর ( ঐচ্ছিক)
মডেলটি যে পরিমাণ চিন্তার টোকেন তৈরি করবে।
সম্ভাব্য মান
-
minimalখুব কম বা একেবারেই চিন্তা না করা।
-
lowচিন্তার নিম্ন স্তর।
-
mediumমাঝারি চিন্তার স্তর।
-
highউচ্চ চিন্তার স্তর।
চিন্তার সারসংক্ষেপ (ঐচ্ছিক)
উত্তরে চিন্তার সারাংশ অন্তর্ভুক্ত করা হবে কিনা।
সম্ভাব্য মান
-
autoস্বয়ংক্রিয় চিন্তার সারাংশ।
-
noneচিন্তামূলক সারসংক্ষেপ নয়।
টুল পছন্দের কনফিগারেশন।
সম্ভাব্য মানসমূহ:
-
autoস্বয়ংক্রিয় টুল নির্বাচন।
-
anyযেকোনো সরঞ্জাম পছন্দ।
-
noneসরঞ্জাম পছন্দের কোনো সুযোগ নেই।
-
validatedসরঞ্জামটির নির্বাচন যথার্থ প্রমাণিত।
ট্রান্সক্রিপশন_কনফিগ ট্রান্সক্রিপশনকনফিগ (ঐচ্ছিক)
ঐচ্ছিক। স্পিচ রিকগনিশন (ট্রান্সক্রিপশন)-এর জন্য কনফিগারেশন। এটি উপস্থিত থাকলে, ASR সক্রিয় হয়।
ক্ষেত্র
ঐচ্ছিক। নির্দিষ্ট পদ শনাক্ত করার দিকে স্পিচ রিকগনিশন মডেলকে প্রভাবিত করার জন্য নিজস্ব শব্দভাণ্ডারের একটি তালিকা।
ঐচ্ছিক। BCP-47 ভাষা কোড যা অডিওতে উপস্থিত ভাষাগুলো সম্পর্কে ইঙ্গিত দেয়। এটি বাদ দিলে বা খালি রাখলে, ডিফল্টরূপে স্বয়ংক্রিয় ভাষা শনাক্তকরণ ব্যবহৃত হবে।
মোড ট্রান্সক্রিপশনমোড অথবা এনাম (স্ট্রিং) (ঐচ্ছিক)
পৃথকীকৃত ট্রান্সক্রিপশন মোড বিকল্প বা এনাম।
সম্ভাব্য প্রকার
স্মার্ট ট্রান্সক্রিপশন মোড
স্মার্ট ট্রান্সক্রিপশন মোডের জন্য কনফিগারেশন।
কোনো বিবরণ দেওয়া হয়নি।
সর্বদা "smart" মোডে সেট করা থাকে।
ভার্ব্যাটিম ট্রান্সক্রিপশন মোড
হুবহু প্রতিলিপি মোডের জন্য কনফিগারেশন।
ঐচ্ছিক। স্পিকার ডায়ারাইজেশন কনফিগার করে। সমর্থিত মান: "speaker"।
ঐচ্ছিক। ট্রান্সক্রিপশন আউটপুটে অন্তর্ভুক্ত করার জন্য টাইমস্ট্যাম্পের সূক্ষ্মতা। সমর্থিত মান: 'word'। খালি থাকলে, কোনো টাইমস্ট্যাম্প তৈরি হবে না।
কোনো বিবরণ দেওয়া হয়নি।
সর্বদা "verbatim" হিসেবে সেট করা থাকে।
ভিডিও_কনফিগ ভিডিওকনফিগ (ঐচ্ছিক)
ভিডিও তৈরির জন্য কনফিগারেশন।
ক্ষেত্র
ভিডিও তৈরির জন্য ঐচ্ছিক টাস্ক মোড। এটি নির্দিষ্ট করা না থাকলে, মডেলটি প্রদত্ত টেক্সট প্রম্পট এবং ইনপুট মিডিয়ার উপর ভিত্তি করে স্বয়ংক্রিয়ভাবে উপযুক্ত মোড নির্ধারণ করে।
সম্ভাব্য মানসমূহ:
-
text_to_videoশুধুমাত্র টেক্সট প্রম্পট থেকে ভিডিও তৈরি করে।
-
image_to_videoএক বা দুটি উৎস ছবি থেকে ভিডিও তৈরি করে। প্রথম ছবিটি শুরুর ফ্রেম এবং ঐচ্ছিক দ্বিতীয় ছবিটি শেষের ফ্রেম নির্ধারণ করে।
-
reference_to_videoরেফারেন্স মিডিয়া (যেমন ছবি, অডিও বা ভিডিও) ব্যবহার করে ভিডিও তৈরি করে।
-
editবিদ্যমান ইনপুট ভিডিও পরিবর্তন করে।
-
extendবিদ্যমান ইনপুট ভিডিওকে সম্প্রসারিত করে।
agent_config অবজেক্ট (ঐচ্ছিক)
এজেন্ট কনফিগারেশন
এজেন্টের জন্য কনফিগারেশন।
`generation_config`-এর বিকল্প। শুধুমাত্র তখনই প্রযোজ্য যখন `agent` সেট করা থাকে।
সম্ভাব্য প্রকার
পলিমরফিক ডিসক্রিমিনেটর: type
AntigravityAgentConfig
অ্যান্টিগ্র্যাভিটি এজেন্ট রানটাইমের কনফিগারেশন। এটি এজেন্টের এক্সিকিউশন এনভায়রনমেন্ট এবং টুল কনফিগারেশনের ওপর সার্ভার-সাইড নিয়ন্ত্রণ প্রদান করে।
এজেন্ট রানের জন্য সর্বোচ্চ মোট টোকেন।
এজেন্টের যুক্তিবোধের জন্য ব্যবহারযোগ্য মডেল।
কোনো বিবরণ দেওয়া হয়নি।
সর্বদা "antigravity" তে সেট করা থাকে।
কোডমেন্ডারএজেন্টকনফিগ
কোডমেন্ডার এজেন্টের কনফিগারেশন।
find_request FindRequest (ঐচ্ছিক)
দুর্বলতা খুঁজে বের করার মাপকাঠি।
ক্ষেত্র
দুর্বলতা বিশ্লেষণকে নির্দেশনা দেওয়ার জন্য ব্যবহারকারী কর্তৃক প্রদত্ত অতিরিক্ত প্রাসঙ্গিক তথ্য বা নিজস্ব নির্দেশাবলী।
যাচাই করার জন্য একটি নির্দিষ্ট অনুসন্ধানের শনাক্তকারী। এটি প্রধানত ভেরিফাই মোডে ব্যবহৃত হয়, যাতে এজেন্টের কার্যসম্পাদন-ভিত্তিক যাচাইকরণ একটিমাত্র দুর্বলতার উপর কেন্দ্রীভূত হয়।
অনুসন্ধান সেশনের মোড।
সম্ভাব্য মানসমূহ:
-
scanশুধুমাত্র প্রাথমিক ক্লাসিফায়ার ব্যবহার করে দ্রুত স্ক্যান।
-
verifyশ্রেণিবিন্যাস সম্পাদনের পর বিস্তারিত তদন্ত করা হয়।
উৎস_ফাইল অ্যারে (ফাইলের বিষয়বস্তু) (ঐচ্ছিক)
স্ক্যানের প্রেক্ষাপট হিসেবে সরবরাহ করার জন্য উৎস ফাইলগুলির একটি তালিকা।
ক্ষেত্র
ফাইলটির UTF-8 এনকোডেড টেক্সট কন্টেন্ট।
প্রজেক্ট রুট থেকে ফাইলটির আপেক্ষিক পথ।
fix_request FixRequest (ঐচ্ছিক)
দুর্বলতা নিরাময়ের জন্য প্যারামিটারসমূহ।
ক্ষেত্র
প্যাচ তৈরির প্রক্রিয়াকে নির্দেশনা দেওয়ার জন্য ব্যবহারকারী কর্তৃক প্রদত্ত অতিরিক্ত প্রাসঙ্গিক তথ্য বা নিজস্ব নির্দেশাবলী।
যে নির্দিষ্ট নিরাপত্তা ত্রুটিটির প্রতিকার করা হবে, তার শনাক্তকারী। এই আইডিটি পূর্বে আবিষ্কৃত একটি দুর্বলতার সাথে সম্পর্কিত।
উৎস_ফাইল অ্যারে (ফাইলের বিষয়বস্তু) (ঐচ্ছিক)
প্রতিকারের প্রেক্ষাপট প্রদানকারী উৎস ফাইলগুলির একটি তালিকা। এই ফাইলগুলিতেই সাধারণত চিহ্নিত দুর্বলতাটি থাকে।
ক্ষেত্র
ফাইলটির UTF-8 এনকোডেড টেক্সট কন্টেন্ট।
প্রজেক্ট রুট থেকে ফাইলটির আপেক্ষিক পথ।
কোডমেন্ডার এজেন্টের জন্য ব্যবহৃত মডেলের নাম। একটি কোডমেন্ডার সেশনে শুধুমাত্র একটি মডেল ব্যবহৃত হবে।
সেশন_কনফিগ সেশনকনফিগ (ঐচ্ছিক)
ডিফল্ট এজেন্ট আচরণকে অগ্রাহ্য করার জন্য ঐচ্ছিক সেশন-নির্দিষ্ট কনফিগারেশন।
ক্ষেত্র
টাইমআউটে পৌঁছানোর আগে এজেন্টকে সর্বাধিক যতগুলো ইন্টারঅ্যাকশন রাউন্ড সম্পাদন করার অনুমতি দেওয়া হয়।
একই CodeMender সেশনের অন্তর্গত একাধিক ইন্টারঅ্যাকশনকে একত্রিত করার প্যারামিটার।
কোনো বিবরণ দেওয়া হয়নি।
সর্বদা "code-mender" এ সেট করা থাকে।
DeepResearchAgentConfig
ডিপ রিসার্চ এজেন্টের কনফিগারেশন।
ডিপ রিসার্চ এজেন্টের জন্য মানব-সম্পৃক্ত পরিকল্পনা সক্ষম করে। যদি এটি 'true' সেট করা হয়, তাহলে ডিপ রিসার্চ এজেন্ট তার প্রতিক্রিয়ায় একটি গবেষণা পরিকল্পনা প্রদান করবে। এরপর এজেন্টটি কেবল তখনই অগ্রসর হবে, যদি ব্যবহারকারী পরবর্তী টার্নে পরিকল্পনাটি নিশ্চিত করে।
ডিপ রিসার্চ এজেন্টের জন্য বিগকোয়েরি টুল সক্রিয় করে।
চিন্তার সারসংক্ষেপ (ঐচ্ছিক)
উত্তরে চিন্তার সারাংশ অন্তর্ভুক্ত করা হবে কিনা।
সম্ভাব্য মান
-
autoস্বয়ংক্রিয় চিন্তার সারাংশ।
-
noneচিন্তামূলক সারসংক্ষেপ নয়।
কোনো বিবরণ দেওয়া হয়নি।
সর্বদা "deep-research" তে সেট করা থাকে।
উত্তরে ভিজ্যুয়ালাইজেশন অন্তর্ভুক্ত করা হবে কিনা।
সম্ভাব্য মানসমূহ:
-
offভিজ্যুয়ালাইজেশন অন্তর্ভুক্ত করবেন না।
-
autoস্বয়ংক্রিয়ভাবে ভিজ্যুয়ালাইজেশন অন্তর্ভুক্ত করুন।
ডাইনামিকএজেন্টকনফিগ
ডাইনামিক এজেন্টদের জন্য কনফিগারেশন।
কোনো বিবরণ দেওয়া হয়নি।
সর্বদা "dynamic" এ সেট করা থাকে।
ইন্টারঅ্যাকশনের জন্য পরিবেশ কনফিগারেশন। এটি রিমোট এনভায়রনমেন্ট সোর্স নির্দিষ্টকারী একটি অবজেক্ট অথবা বিদ্যমান এনভায়রনমেন্ট আইডি উল্লেখকারী একটি স্ট্রিং হতে পারে।
অনুরোধটির জন্য ব্যবহারকারী-সংজ্ঞায়িত মেটাডেটা সহ লেবেলগুলি।
পূর্ববর্তী যোগাযোগের আইডি, যদি থাকে।
safety_settings অ্যারে (নিরাপত্তা সেটিংস) (ঐচ্ছিক)
মিথস্ক্রিয়ার জন্য নিরাপত্তা সেটিংস।
ক্ষেত্র
ঐচ্ছিক। কন্টেন্ট ব্লক করার পদ্ধতি। নির্দিষ্ট করে না দিলে, ডিফল্ট আচরণ হিসেবে সম্ভাব্যতা স্কোর ব্যবহার করা হয়।
সম্ভাব্য মানসমূহ:
-
severityক্ষতি ব্লক পদ্ধতিতে সম্ভাবনা এবং তীব্রতা উভয় স্কোরই ব্যবহার করা হয়।
-
probabilityক্ষতি ব্লক পদ্ধতিটি সম্ভাব্যতা স্কোর ব্যবহার করে।
আবশ্যক। কন্টেন্ট ব্লক করার সীমা। ক্ষতির সম্ভাবনা এই সীমা অতিক্রম করলে কন্টেন্টটি ব্লক করা হবে।
সম্ভাব্য মানসমূহ:
-
block_low_and_aboveকম বা তার বেশি ক্ষতির সম্ভাবনা আছে এমন কন্টেন্ট ব্লক করুন।
-
block_medium_and_aboveমাঝারি বা তার চেয়ে বেশি ক্ষতির সম্ভাবনা রয়েছে এমন কন্টেন্ট ব্লক করুন।
-
block_only_highক্ষতির উচ্চ সম্ভাবনাযুক্ত বিষয়বস্তু ব্লক করুন।
-
block_noneক্ষতির সম্ভাবনা নির্বিশেষে কোনো বিষয়বস্তু ব্লক করবেন না।
-
offসুরক্ষা ফিল্টারটি পুরোপুরি বন্ধ করে দিন।
ক্ষতির বিভাগ ( ঐচ্ছিক)
আবশ্যক। যে ধরনের ক্ষতির বিভাগ অবরুদ্ধ করতে হবে।
সম্ভাব্য মান
-
hate_speechএমন বিষয়বস্তু যা নির্দিষ্ট বৈশিষ্ট্যের ভিত্তিতে কোনো ব্যক্তি বা গোষ্ঠীর বিরুদ্ধে সহিংসতাকে উৎসাহিত করে বা ঘৃণা উস্কে দেয়।
-
dangerous_contentএমন বিষয়বস্তু যা বিপজ্জনক কার্যকলাপকে উৎসাহিত করে, সহজতর করে বা সক্ষম করে তোলে।
-
harassmentঅপমানজনক, হুমকিমূলক, অথবা উৎপীড়ন, নির্যাতন বা উপহাস করার উদ্দেশ্যে তৈরি বিষয়বস্তু।
-
sexually_explicitযেসব বিষয়বস্তুতে যৌনতাপূর্ণ উপাদান রয়েছে।
-
civic_integrityঅপ্রচলিত: নির্বাচন ফিল্টার আর সমর্থিত নয়। ক্ষতির বিভাগটি হলো নাগরিক অখণ্ডতা।
-
image_hateবিদ্বেষমূলক বক্তব্য সম্বলিত ছবি।
-
image_dangerous_contentযেসব ছবিতে বিপজ্জনক বিষয়বস্তু রয়েছে।
-
image_harassmentযেসব ছবিতে হয়রানি রয়েছে।
-
image_sexually_explicitযেসব ছবিতে যৌন উত্তেজক বিষয়বস্তু রয়েছে।
-
jailbreakসুরক্ষা ফিল্টার এড়িয়ে যাওয়ার জন্য ডিজাইন করা প্রম্পট।
সার্ভিস_টিয়ার সার্ভিসটিয়ার (ঐচ্ছিক)
মিথস্ক্রিয়ার জন্য পরিষেবা স্তর।
সম্ভাব্য মান
-
flexনমনীয় পরিষেবা স্তর।
-
standardসাধারণ পরিষেবা স্তর।
-
priorityঅগ্রাধিকার পরিষেবা স্তর।
-
deferredবিলম্বিত পরিষেবা স্তর।
webhook_config WebhookConfig (ঐচ্ছিক)
ঐচ্ছিক। ইন্টারঅ্যাকশন সম্পন্ন হলে নোটিফিকেশন পাওয়ার জন্য ওয়েবহুক কনফিগারেশন।
ক্ষেত্র
ঐচ্ছিক। সেট করা হলে, নিবন্ধিত ওয়েবহুকগুলোর পরিবর্তে এই ওয়েবহুক ইউআরআইগুলো ওয়েবহুক ইভেন্টের জন্য ব্যবহৃত হবে।
ঐচ্ছিক। ব্যবহারকারীর মেটাডেটা যা ওয়েবহুকগুলিতে প্রতিটি ইভেন্ট প্রেরণের সময় ফেরত দেওয়া হবে।
প্রতিক্রিয়া
একটি ইন্টারঅ্যাকশন রিসোর্স ফেরত দেয়।
সাধারণ অনুরোধ
উদাহরণ প্রতিক্রিয়া
{ "created": "2025-11-26T12:25:15Z", "id": "v1_ChdPU0F4YWFtNkFwS2kxZThQZ05lbXdROBIXT1NBeGFhbTZBcEtpMWU4UGdOZW13UTg", "model": "gemini-3.6-flash", "object": "interaction", "status": "completed", "steps": [ { "type": "model_output", "content": [ { "type": "text", "text": "Hello! I'm functioning perfectly and ready to assist you.\n\nHow are you doing today?" } ] } ], "updated": "2025-11-26T12:25:15Z", "usage": { "input_tokens_by_modality": [ { "modality": "text", "tokens": 7 } ], "total_cached_tokens": 0, "total_input_tokens": 7, "total_output_tokens": 20, "total_thought_tokens": 22, "total_tokens": 49, "total_tool_use_tokens": 0 } }
মাল্টি-টার্ন
উদাহরণ প্রতিক্রিয়া
{ "created": "2025-11-26T12:22:47Z", "id": "v1_ChdPU0F4YWFtNkFwS2kxZThQZ05lbXdROBIXT1NBeGFhbTZBcEtpMWU4UGdOZW13UTg", "model": "gemini-3.6-flash", "object": "interaction", "status": "completed", "steps": [ { "type": "model_output", "content": [ { "type": "text", "text": "The capital of France is Paris." } ] } ], "updated": "2025-11-26T12:22:47Z", "usage": { "input_tokens_by_modality": [ { "modality": "text", "tokens": 50 } ], "total_cached_tokens": 0, "total_input_tokens": 50, "total_output_tokens": 10, "total_thought_tokens": 0, "total_tokens": 60, "total_tool_use_tokens": 0 } }
ইমেজ ইনপুট
উদাহরণ প্রতিক্রিয়া
{ "created": "2025-11-26T12:22:47Z", "id": "v1_ChdPU0F4YWFtNkFwS2kxZThQZ05lbXdROBIXT1NBeGFhbTZBcEtpMWU4UGdOZW13UTg", "model": "gemini-3.6-flash", "object": "interaction", "status": "completed", "steps": [ { "type": "model_output", "content": [ { "type": "text", "text": "A white humanoid robot with glowing blue eyes stands holding a red skateboard." } ] } ], "updated": "2025-11-26T12:22:47Z", "usage": { "input_tokens_by_modality": [ { "modality": "text", "tokens": 10 }, { "modality": "image", "tokens": 258 } ], "total_cached_tokens": 0, "total_input_tokens": 268, "total_output_tokens": 20, "total_thought_tokens": 0, "total_tokens": 288, "total_tool_use_tokens": 0 } }
ফাংশন কলিং
উদাহরণ প্রতিক্রিয়া
{ "created": "2025-11-26T12:22:47Z", "id": "v1_ChdPU0F4YWFtNkFwS2kxZThQZ05lbXdROBIXT1NBeGFhbTZBcEtpMWU4UGdOZW13UTg", "model": "gemini-3.6-flash", "object": "interaction", "status": "requires_action", "steps": [ { "name": "get_weather", "type": "function_call", "arguments": { "location": "Boston, MA" }, "id": "gth23981" } ], "updated": "2025-11-26T12:22:47Z", "usage": { "input_tokens_by_modality": [ { "modality": "text", "tokens": 100 } ], "total_cached_tokens": 0, "total_input_tokens": 100, "total_output_tokens": 25, "total_thought_tokens": 0, "total_tokens": 125, "total_tool_use_tokens": 50 } }
গভীর গবেষণা
উদাহরণ প্রতিক্রিয়া
{ "agent": "deep-research-pro-preview-12-2025", "created": "2025-11-26T12:22:47Z", "id": "v1_ChdPU0F4YWFtNkFwS2kxZThQZ05lbXdROBIXT1NBeGFhbTZBcEtpMWU4UGdOZW13UTg", "object": "interaction", "status": "completed", "steps": [ { "type": "model_output", "content": [ { "type": "text", "text": "Here is a comprehensive research report on the current state of cancer research..." } ] } ], "updated": "2025-11-26T12:22:47Z", "usage": { "input_tokens_by_modality": [ { "modality": "text", "tokens": 20 } ], "total_cached_tokens": 0, "total_input_tokens": 20, "total_output_tokens": 1000, "total_thought_tokens": 500, "total_tokens": 1520, "total_tool_use_tokens": 0 } }
অ্যান্টিগ্র্যাভিটি এজেন্ট
উদাহরণ প্রতিক্রিয়া
{ "agent": "antigravity-preview-05-2026", "created": "2025-11-26T12:22:47Z", "environment_id": "env_abc123", "id": "v1_ChdPU0F4YWFtNkFwS2kxZThQZ05lbXdROBIXT1NBeGFhbTZBcEtpMWU4UGdOZW13UTg", "object": "interaction", "status": "completed", "steps": [ { "type": "model_output", "content": [ { "type": "text", "text": "I've summarized the top 5 Hacker News stories and saved the results to /workspace/summary.md." } ] } ], "updated": "2025-11-26T12:22:47Z", "usage": { "input_tokens_by_modality": [ { "modality": "text", "tokens": 50 } ], "total_cached_tokens": 0, "total_input_tokens": 50, "total_output_tokens": 500, "total_thought_tokens": 200, "total_tokens": 750, "total_tool_use_tokens": 0 } }
পুনর্ব্যবহার পরিবেশ
উদাহরণ প্রতিক্রিয়া
{ "agent": "antigravity-preview-05-2026", "created": "2025-11-26T12:23:00Z", "environment_id": "env_abc123", "id": "v1_Chd2ZTJhYmNkZWZnaGlqa2xtbm9wcXJzdHV2d3h5ejAxMjM0NTY3ODkwMTIzNDU2Nzg", "object": "interaction", "status": "completed", "steps": [ { "type": "model_output", "content": [ { "type": "text", "text": "I've updated /workspace/hello.py to accept a name argument and greet the user." } ] } ], "updated": "2025-11-26T12:23:00Z", "usage": { "input_tokens_by_modality": [ { "modality": "text", "tokens": 80 } ], "total_cached_tokens": 0, "total_input_tokens": 80, "total_output_tokens": 200, "total_thought_tokens": 100, "total_tokens": 380, "total_tool_use_tokens": 0 } }
সূত্র সহ
কাস্টম এজেন্ট
একটি ইন্টারঅ্যাকশন বাতিল করা
আইডি দ্বারা একটি ইন্টারঅ্যাকশন বাতিল করে। এটি শুধুমাত্র চলমান ব্যাকগ্রাউন্ড ইন্টারঅ্যাকশনগুলোর ক্ষেত্রে প্রযোজ্য।
পাথ / কোয়েরি প্যারামিটার
এপিআই-এর কোন সংস্করণটি ব্যবহার করতে হবে।
বাতিল করার জন্য ইন্টারঅ্যাকশনটির অনন্য শনাক্তকারী।
প্রতিক্রিয়া
একটি ইন্টারঅ্যাকশন রিসোর্স ফেরত দেয়।
মিথস্ক্রিয়া বাতিল করুন
উদাহরণ প্রতিক্রিয়া
{ "agent": "deep-research-pro-preview-12-2025", "created": "2026-06-22T04:55:47Z", "id": "v1_ChdVc0E0YXJTYk1zYlV6N0lQcXRXVG1BYxIXVXNBNGFyU2JNc2JVejdJUHF0V1RtQWM", "status": "cancelled", "steps": [ { "type": "user_input", "content": [ { "type": "text", "text": "Research the history of the Google TPUs with a focus on 2025 specs." } ] } ], "updated": "2026-06-22T04:55:47Z" }
একটি মিথস্ক্রিয়া পুনরুদ্ধার করা
`Interaction.id`-এর উপর ভিত্তি করে একটিমাত্র ইন্টারঅ্যাকশনের সম্পূর্ণ বিবরণ পুনরুদ্ধার করে।
পাথ / কোয়েরি প্যারামিটার
এপিআই-এর কোন সংস্করণটি ব্যবহার করতে হবে।
পুনরুদ্ধার করার জন্য ইন্টারঅ্যাকশনটির অনন্য শনাক্তকারী।
ঐচ্ছিক। সেট করা থাকলে, ইভেন্ট আইডি দ্বারা চিহ্নিত ইভেন্টের পরের চাঙ্ক থেকে ইন্টারঅ্যাকশন স্ট্রিম পুনরায় শুরু হয়। এটি শুধুমাত্র তখনই ব্যবহার করা যাবে যখন `stream` সত্য হবে।
true-তে সেট করা হলে, তৈরি হওয়া কন্টেন্ট পর্যায়ক্রমে স্ট্রিম করা হবে।
ডিফল্ট মান: False
প্রতিক্রিয়া
একটি ইন্টারঅ্যাকশন রিসোর্স ফেরত দেয়।
মিথস্ক্রিয়া করুন
উদাহরণ প্রতিক্রিয়া
{ "created": "2025-11-26T12:25:15Z", "id": "v1_ChdPU0F4YWFtNkFwS2kxZThQZ05lbXdROBIXT1NBeGFhbTZBcEtpMWU4UGdOZW13UTg", "model": "gemini-3.6-flash", "object": "interaction", "status": "completed", "steps": [ { "type": "model_output", "content": [ { "type": "text", "text": "I'm doing great, thank you for asking! How can I help you today?" } ] } ], "updated": "2025-11-26T12:25:15Z" }
একটি ইন্টারঅ্যাকশন মুছে ফেলা
আইডি দ্বারা ইন্টারঅ্যাকশনটি মুছে দেয়।
পাথ / কোয়েরি প্যারামিটার
এপিআই-এর কোন সংস্করণটি ব্যবহার করতে হবে।
মুছে ফেলার জন্য ইন্টারঅ্যাকশনটির অনন্য শনাক্তকারী।
প্রতিক্রিয়া
সফল হলে, প্রতিক্রিয়াটি খালি থাকে।
মুছে ফেলুন
সম্পদ
মিথস্ক্রিয়া
মিথস্ক্রিয়া সম্পদ।
ক্ষেত্র
এজেন্ট এজেন্টঅপশন (ঐচ্ছিক)
ইন্টারঅ্যাকশনটি তৈরি করতে ব্যবহৃত 'এজেন্ট'-এর নাম।
সম্ভাব্য মান
-
deep-research-pro-preview-12-2025জেমিনি ডিপ রিসার্চ এজেন্ট
-
deep-research-preview-04-2026জেমিনি ডিপ রিসার্চ এজেন্ট
-
deep-research-max-preview-04-2026জেমিনি ডিপ রিসার্চ ম্যাক্স এজেন্ট
-
antigravity-preview-05-2026যুক্তিবোধ, ফাইল পরিচালনা এবং টুল ব্যবহারের প্রয়োজন হয় এমন একাধিক ধাপের কাজ সম্পাদন করতে অ্যান্টিগ্র্যাভিটি পরিচালিত এজেন্টটি ব্যবহার করুন।
agent_config অবজেক্ট (ঐচ্ছিক)
এজেন্ট ইন্টারঅ্যাকশনের জন্য কনফিগারেশন প্যারামিটারসমূহ।
সম্ভাব্য প্রকার
পলিমরফিক ডিসক্রিমিনেটর: type
AntigravityAgentConfig
অ্যান্টিগ্র্যাভিটি এজেন্ট রানটাইমের কনফিগারেশন। এটি এজেন্টের এক্সিকিউশন এনভায়রনমেন্ট এবং টুল কনফিগারেশনের ওপর সার্ভার-সাইড নিয়ন্ত্রণ প্রদান করে।
এজেন্ট রানের জন্য সর্বোচ্চ মোট টোকেন।
এজেন্টের যুক্তিবোধের জন্য ব্যবহারযোগ্য মডেল।
কোনো বিবরণ দেওয়া হয়নি।
সর্বদা "antigravity" তে সেট করা থাকে।
কোডমেন্ডারএজেন্টকনফিগ
কোডমেন্ডার এজেন্টের কনফিগারেশন।
find_request FindRequest (ঐচ্ছিক)
দুর্বলতা খুঁজে বের করার মাপকাঠি।
ক্ষেত্র
দুর্বলতা বিশ্লেষণকে নির্দেশনা দেওয়ার জন্য ব্যবহারকারী কর্তৃক প্রদত্ত অতিরিক্ত প্রাসঙ্গিক তথ্য বা নিজস্ব নির্দেশাবলী।
যাচাই করার জন্য একটি নির্দিষ্ট অনুসন্ধানের শনাক্তকারী। এটি প্রধানত ভেরিফাই মোডে ব্যবহৃত হয়, যাতে এজেন্টের কার্যসম্পাদন-ভিত্তিক যাচাইকরণ একটিমাত্র দুর্বলতার উপর কেন্দ্রীভূত হয়।
অনুসন্ধান সেশনের মোড।
সম্ভাব্য মানসমূহ:
-
scanশুধুমাত্র প্রাথমিক ক্লাসিফায়ার ব্যবহার করে দ্রুত স্ক্যান।
-
verifyশ্রেণিবিন্যাস সম্পাদনের পর বিস্তারিত তদন্ত করা হয়।
উৎস_ফাইল অ্যারে (ফাইলের বিষয়বস্তু) (ঐচ্ছিক)
স্ক্যানের প্রেক্ষাপট হিসেবে সরবরাহ করার জন্য উৎস ফাইলগুলির একটি তালিকা।
ক্ষেত্র
ফাইলটির UTF-8 এনকোডেড টেক্সট কন্টেন্ট।
প্রজেক্ট রুট থেকে ফাইলটির আপেক্ষিক পথ।
fix_request FixRequest (ঐচ্ছিক)
দুর্বলতা নিরাময়ের জন্য প্যারামিটারসমূহ।
ক্ষেত্র
প্যাচ তৈরির প্রক্রিয়াকে নির্দেশনা দেওয়ার জন্য ব্যবহারকারী কর্তৃক প্রদত্ত অতিরিক্ত প্রাসঙ্গিক তথ্য বা নিজস্ব নির্দেশাবলী।
যে নির্দিষ্ট নিরাপত্তা ত্রুটিটির প্রতিকার করা হবে, তার শনাক্তকারী। এই আইডিটি পূর্বে আবিষ্কৃত একটি দুর্বলতার সাথে সম্পর্কিত।
উৎস_ফাইল অ্যারে (ফাইলের বিষয়বস্তু) (ঐচ্ছিক)
প্রতিকারের প্রেক্ষাপট প্রদানকারী উৎস ফাইলগুলির একটি তালিকা। এই ফাইলগুলিতেই সাধারণত চিহ্নিত দুর্বলতাটি থাকে।
ক্ষেত্র
ফাইলটির UTF-8 এনকোডেড টেক্সট কন্টেন্ট।
প্রজেক্ট রুট থেকে ফাইলটির আপেক্ষিক পথ।
কোডমেন্ডার এজেন্টের জন্য ব্যবহৃত মডেলের নাম। একটি কোডমেন্ডার সেশনে শুধুমাত্র একটি মডেল ব্যবহৃত হবে।
সেশন_কনফিগ সেশনকনফিগ (ঐচ্ছিক)
ডিফল্ট এজেন্ট আচরণকে অগ্রাহ্য করার জন্য ঐচ্ছিক সেশন-নির্দিষ্ট কনফিগারেশন।
ক্ষেত্র
টাইমআউটে পৌঁছানোর আগে এজেন্টকে সর্বাধিক যতগুলো ইন্টারঅ্যাকশন রাউন্ড সম্পাদন করার অনুমতি দেওয়া হয়।
একই CodeMender সেশনের অন্তর্গত একাধিক ইন্টারঅ্যাকশনকে একত্রিত করার প্যারামিটার।
কোনো বিবরণ দেওয়া হয়নি।
সর্বদা "code-mender" এ সেট করা থাকে।
DeepResearchAgentConfig
ডিপ রিসার্চ এজেন্টের কনফিগারেশন।
ডিপ রিসার্চ এজেন্টের জন্য মানব-সম্পৃক্ত পরিকল্পনা সক্ষম করে। যদি এটি 'true' সেট করা হয়, তাহলে ডিপ রিসার্চ এজেন্ট তার প্রতিক্রিয়ায় একটি গবেষণা পরিকল্পনা প্রদান করবে। এরপর এজেন্টটি কেবল তখনই অগ্রসর হবে, যদি ব্যবহারকারী পরবর্তী টার্নে পরিকল্পনাটি নিশ্চিত করে।
ডিপ রিসার্চ এজেন্টের জন্য বিগকোয়েরি টুল সক্রিয় করে।
চিন্তার সারসংক্ষেপ (ঐচ্ছিক)
উত্তরে চিন্তার সারাংশ অন্তর্ভুক্ত করা হবে কিনা।
সম্ভাব্য মান
-
autoস্বয়ংক্রিয় চিন্তার সারাংশ।
-
noneচিন্তামূলক সারসংক্ষেপ নয়।
কোনো বিবরণ দেওয়া হয়নি।
সর্বদা "deep-research" তে সেট করা থাকে।
উত্তরে ভিজ্যুয়ালাইজেশন অন্তর্ভুক্ত করা হবে কিনা।
সম্ভাব্য মানসমূহ:
-
offভিজ্যুয়ালাইজেশন অন্তর্ভুক্ত করবেন না।
-
autoস্বয়ংক্রিয়ভাবে ভিজ্যুয়ালাইজেশন অন্তর্ভুক্ত করুন।
ডাইনামিকএজেন্টকনফিগ
ডাইনামিক এজেন্টদের জন্য কনফিগারেশন।
কোনো বিবরণ দেওয়া হয়নি।
সর্বদা "dynamic" এ সেট করা থাকে।
শুধুমাত্র আউটপুট। যে সময়ে প্রতিক্রিয়াটি তৈরি করা হয়েছিল, সেই সময়টি ISO 8601 ফরম্যাটে (YYYY-MM-DDThh:mm:ssZ) উল্লেখ করতে হবে।
ইন্টারঅ্যাকশনের জন্য পরিবেশ কনফিগারেশন। এটি রিমোট এনভায়রনমেন্ট সোর্স নির্দিষ্টকারী একটি অবজেক্ট অথবা বিদ্যমান এনভায়রনমেন্ট আইডি উল্লেখকারী একটি স্ট্রিং হতে পারে।
শুধুমাত্র আউটপুট। ইন্টারঅ্যাকশনটির জন্য এনভায়রনমেন্ট আইডি। অনুরোধে এনভায়রনমেন্ট কনফিগারেশন সেট করা থাকলেই এটি পূরণ করা হবে।
ত্রুটি অ্যারে (ত্রুটি) (ঐচ্ছিক)
শুধুমাত্র আউটপুট। ইন্টারঅ্যাকশনে ডায়াগনস্টিক ত্রুটি / প্ল্যাটফর্ম ভুল রেকর্ড করা হয়েছে।
ক্ষেত্র
একটি URI যা ত্রুটির ধরণ শনাক্ত করে।
মানুষের পাঠযোগ্য একটি ত্রুটি বার্তা।
আবশ্যক। শুধুমাত্র আউটপুট। মিথস্ক্রিয়া সম্পন্ন করার জন্য একটি অনন্য শনাক্তকারী।
ডিফল্ট হলো:
মিথস্ক্রিয়ার জন্য ইনপুট।
অনুরোধটির জন্য ব্যবহারকারী-সংজ্ঞায়িত মেটাডেটা সহ লেবেলগুলি।
মডেল মডেলঅপশন (ঐচ্ছিক)
ইন্টারঅ্যাকশনটি তৈরি করতে ব্যবহৃত `মডেল`-এর নাম।
সম্ভাব্য মান
-
gemini-2.5-flashআমাদের প্রথম হাইব্রিড রিজনিং মডেল যা ১ মিলিয়ন টোকেন কনটেক্সট উইন্ডো এবং থিংকিং বাজেট সমর্থন করে।
-
gemini-2.5-proআমাদের সর্বাধুনিক বহুমুখী মডেল, যা কোডিং এবং জটিল যুক্তিনির্ভর কাজে অত্যন্ত পারদর্শী।
-
gemma-4-26b-a4b-itজেমা 4 26B A4B IT
-
gemma-4-31b-itজেমা ৪ ৩১বি আইটি
-
gemini-flash-latestজেমিনি ফ্ল্যাশের সর্বশেষ সংস্করণ
-
gemini-flash-lite-latestজেমিনি ফ্ল্যাশ-লাইটের সর্বশেষ সংস্করণ
-
gemini-pro-latestজেমিনি প্রো-এর সর্বশেষ সংস্করণ
-
gemini-2.5-flash-liteআমাদের সবচেয়ে ছোট এবং সবচেয়ে সাশ্রয়ী মডেল, যা ব্যাপক ব্যবহারের জন্য নির্মিত।
-
gemini-2.5-flash-imageআমাদের নিজস্ব ইমেজ জেনারেশন মডেলটি গতি, নমনীয়তা এবং প্রাসঙ্গিকতা বোঝার জন্য অপ্টিমাইজ করা হয়েছে। টেক্সট ইনপুট এবং আউটপুটের মূল্য ২.৫ ফ্ল্যাশের সমান।
-
gemini-3-flash-previewগতির জন্য নির্মিত আমাদের সবচেয়ে বুদ্ধিমান মডেল, যা অত্যাধুনিক বুদ্ধিমত্তার সাথে উন্নত অনুসন্ধান এবং ভূমিতে স্থিতিশীলতার সমন্বয় ঘটায়।
-
gemini-3.1-pro-previewআমাদের সর্বাধুনিক অত্যাধুনিক রিজনিং মডেল, যা অভূতপূর্ব গভীরতা ও সূক্ষ্মতা এবং শক্তিশালী মাল্টিমোডাল আন্ডারস্ট্যান্ডিং ও কোডিং সক্ষমতাসম্পন্ন।
-
gemini-3.1-pro-preview-customtoolsকাস্টম টুল ব্যবহারের জন্য অপ্টিমাইজ করা জেমিনি ৩.১ প্রো প্রিভিউ।
-
gemini-3.1-flash-liteআমাদের সবচেয়ে সাশ্রয়ী মডেল, যা বিপুল পরিমাণ এজেন্টিক কাজ, অনুবাদ এবং সাধারণ ডেটা প্রক্রিয়াকরণের জন্য অপ্টিমাইজ করা হয়েছে।
-
gemini-3-pro-imageজেমিনি ৩ প্রো ইমেজ
-
nano-banana-pro-previewজেমিনি ৩ প্রো ছবির প্রিভিউ
-
gemini-3.1-flash-imageজেমিনি ৩.১ ফ্ল্যাশ চিত্র।
-
gemini-3.5-flashএজেন্টিক এবং কোডিং কার্যক্রমে ধারাবাহিক অগ্রণী পারফরম্যান্সের জন্য আমাদের সবচেয়ে বুদ্ধিমান মডেল।
-
gemini-3.6-flashএজেন্টিক এবং কোডিং কার্যক্রমে ধারাবাহিক অগ্রণী পারফরম্যান্সের জন্য আমাদের সবচেয়ে বুদ্ধিমান মডেল।
-
gemini-3.7-flashএজেন্টিক এবং কোডিং কার্যক্রমে ধারাবাহিক অগ্রণী পারফরম্যান্সের জন্য আমাদের সবচেয়ে বুদ্ধিমান মডেল।
-
lyria-3-clip-previewআমাদের স্বল্প-বিলম্বের সঙ্গীত তৈরির মডেলটি উচ্চ-মানের অডিও ক্লিপ এবং সুনির্দিষ্ট ছন্দ নিয়ন্ত্রণের জন্য অপ্টিমাইজ করা হয়েছে।
-
lyria-3-pro-previewগভীর সুরসৃষ্টিগত বোধসম্পন্ন আমাদের উন্নত, পূর্ণাঙ্গ গান তৈরির জেনারেটিভ মডেলটি, বিভিন্ন সঙ্গীত শৈলীতে সুনির্দিষ্ট কাঠামোগত নিয়ন্ত্রণ এবং জটিল রূপান্তরের জন্য সর্বোত্তমভাবে প্রস্তুত করা হয়েছে।
-
gemini-robotics-er-1.6-previewজেমিনি রোবোটিক্স-ইআর ১.৬ প্রিভিউ
-
gemini-robotics-er-2-previewজেমিনি রোবোটিক্স এমবডিড রিজনিং ২ প্রিভিউ
পূর্ববর্তী যোগাযোগের আইডি, যদি থাকে।
এটি নিশ্চিত করে যে তৈরি হওয়া প্রতিক্রিয়াটি একটি JSON অবজেক্ট হবে যা এই ফিল্ডে নির্দিষ্ট করা JSON স্কিমা মেনে চলে।
safety_settings অ্যারে (নিরাপত্তা সেটিংস) (ঐচ্ছিক)
মিথস্ক্রিয়ার জন্য নিরাপত্তা সেটিংস।
ক্ষেত্র
ঐচ্ছিক। কন্টেন্ট ব্লক করার পদ্ধতি। নির্দিষ্ট করে না দিলে, ডিফল্ট আচরণ হিসেবে সম্ভাব্যতা স্কোর ব্যবহার করা হয়।
সম্ভাব্য মানসমূহ:
-
severityক্ষতি ব্লক পদ্ধতিতে সম্ভাবনা এবং তীব্রতা উভয় স্কোরই ব্যবহার করা হয়।
-
probabilityক্ষতি ব্লক পদ্ধতিটি সম্ভাব্যতা স্কোর ব্যবহার করে।
আবশ্যক। কন্টেন্ট ব্লক করার সীমা। ক্ষতির সম্ভাবনা এই সীমা অতিক্রম করলে কন্টেন্টটি ব্লক করা হবে।
সম্ভাব্য মানসমূহ:
-
block_low_and_aboveকম বা তার বেশি ক্ষতির সম্ভাবনা আছে এমন কন্টেন্ট ব্লক করুন।
-
block_medium_and_aboveমাঝারি বা তার চেয়ে বেশি ক্ষতির সম্ভাবনা রয়েছে এমন কন্টেন্ট ব্লক করুন।
-
block_only_highক্ষতির উচ্চ সম্ভাবনাযুক্ত বিষয়বস্তু ব্লক করুন।
-
block_noneক্ষতির সম্ভাবনা নির্বিশেষে কোনো বিষয়বস্তু ব্লক করবেন না।
-
offসুরক্ষা ফিল্টারটি পুরোপুরি বন্ধ করে দিন।
ক্ষতির বিভাগ ( ঐচ্ছিক)
আবশ্যক। যে ধরনের ক্ষতির বিভাগ অবরুদ্ধ করতে হবে।
সম্ভাব্য মান
-
hate_speechএমন বিষয়বস্তু যা নির্দিষ্ট বৈশিষ্ট্যের ভিত্তিতে কোনো ব্যক্তি বা গোষ্ঠীর বিরুদ্ধে সহিংসতাকে উৎসাহিত করে বা ঘৃণা উস্কে দেয়।
-
dangerous_contentএমন বিষয়বস্তু যা বিপজ্জনক কার্যকলাপকে উৎসাহিত করে, সহজতর করে বা সক্ষম করে তোলে।
-
harassmentঅপমানজনক, হুমকিমূলক, অথবা উৎপীড়ন, নির্যাতন বা উপহাস করার উদ্দেশ্যে তৈরি বিষয়বস্তু।
-
sexually_explicitযেসব বিষয়বস্তুতে যৌনতাপূর্ণ উপাদান রয়েছে।
-
civic_integrityঅপ্রচলিত: নির্বাচন ফিল্টার আর সমর্থিত নয়। ক্ষতির বিভাগটি হলো নাগরিক অখণ্ডতা।
-
image_hateবিদ্বেষমূলক বক্তব্য সম্বলিত ছবি।
-
image_dangerous_contentযেসব ছবিতে বিপজ্জনক বিষয়বস্তু রয়েছে।
-
image_harassmentযেসব ছবিতে হয়রানি রয়েছে।
-
image_sexually_explicitযেসব ছবিতে যৌন উত্তেজক বিষয়বস্তু রয়েছে।
-
jailbreakসুরক্ষা ফিল্টার এড়িয়ে যাওয়ার জন্য ডিজাইন করা প্রম্পট।
সার্ভিস_টিয়ার সার্ভিসটিয়ার (ঐচ্ছিক)
মিথস্ক্রিয়ার জন্য পরিষেবা স্তর।
সম্ভাব্য মান
-
flexনমনীয় পরিষেবা স্তর।
-
standardসাধারণ পরিষেবা স্তর।
-
priorityঅগ্রাধিকার পরিষেবা স্তর।
-
deferredবিলম্বিত পরিষেবা স্তর।
আবশ্যক। শুধুমাত্র আউটপুট। মিথস্ক্রিয়ার অবস্থা।
সম্ভাব্য মানসমূহ:
-
in_progressআলোচনাটি চলছে।
-
requires_actionএই মিথস্ক্রিয়ার জন্য ব্যবহারকারীর পক্ষ থেকে কোনো পদক্ষেপ বা ইনপুট প্রয়োজন।
-
completedমিথস্ক্রিয়াটি সম্পন্ন হয়েছে।
-
failedমিথস্ক্রিয়াটি ব্যর্থ হয়েছে।
-
cancelledআলাপচারিতাটি বাতিল করা হয়েছিল।
-
incompleteইন্টারঅ্যাকশনটি সম্পন্ন হয়েছে, কিন্তু এতে অসম্পূর্ণ ফলাফল রয়েছে (যেমন max_tokens-এ পৌঁছানো)।
-
budget_exceededটোকেন বাজেট অতিক্রম করায় মিথস্ক্রিয়াটি স্থগিত করা হয়েছিল।
-
queuedইন্টারঅ্যাকশনটি প্রক্রিয়াকরণের জন্য সারিতে রাখা হয়েছে।
শুধুমাত্র আউটপুট। প্রতিক্রিয়ার অন্তর্ভুক্ত হলে, মিথস্ক্রিয়াটি গঠনকারী ধাপগুলো।
মিথস্ক্রিয়ার জন্য সিস্টেম নির্দেশাবলী।
ইন্টারঅ্যাকশনের সময় মডেলটি যেসব টুল ডিক্লারেশন কল করতে পারে, তার একটি তালিকা।
শুধুমাত্র আউটপুট। যে সময়ে প্রতিক্রিয়াটি সর্বশেষ আপডেট করা হয়েছিল, সেই সময়টি ISO 8601 ফরম্যাটে (YYYY-MM-DDThh:mm:ssZ)।
ব্যবহার (ঐচ্ছিক )
শুধুমাত্র আউটপুট। ইন্টারঅ্যাকশন অনুরোধের টোকেন ব্যবহারের পরিসংখ্যান।
ক্ষেত্র
cached_tokens_by_modality অ্যারে (ModalityTokens) (ঐচ্ছিক)
পদ্ধতি অনুসারে ক্যাশড টোকেন ব্যবহারের বিশদ বিবরণ।
ক্ষেত্র
পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)
টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।
সম্ভাব্য মান
-
textএটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।
-
imageএটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।
-
audioএটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।
-
videoএটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।
-
documentএটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।
মোডালিটির জন্য টোকেনের সংখ্যা।
গ্রাউন্ডিং_টুল_কাউন্ট অ্যারে (GroundingToolCount) (ঐচ্ছিক)
গ্রাউন্ডিং টুলের সংখ্যা।
ক্ষেত্র
গ্রাউন্ডিং টুলের সংখ্যা গুরুত্বপূর্ণ।
গণনার সাথে সংশ্লিষ্ট গ্রাউন্ডিং টুলের ধরণ।
সম্ভাব্য মানসমূহ:
-
google_searchগুগল ওয়েব সার্চ ও ইমেজ সার্চের মাধ্যমে ভিত্তি স্থাপন, এবং এন্টারপ্রাইজের জন্য ওয়েব ভিত্তি স্থাপন।
-
google_mapsগুগল ম্যাপসের সাহায্যে পরিচিতি।
-
retrievalগ্রাহকের ডেটার মাধ্যমে ভিত্তি স্থাপন, উদাহরণস্বরূপ, VertexAISearch।
input_tokens_by_modality অ্যারে (ModalityTokens) (ঐচ্ছিক)
পদ্ধতি অনুসারে ইনপুট টোকেন ব্যবহারের বিশদ বিবরণ।
ক্ষেত্র
পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)
টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।
সম্ভাব্য মান
-
textএটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।
-
imageএটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।
-
audioএটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।
-
videoএটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।
-
documentএটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।
মোডালিটির জন্য টোকেনের সংখ্যা।
আউটপুট_টোকেন_বাই_মোডালিটি অ্যারে (মোডালিটিটোকেন) (ঐচ্ছিক)
পদ্ধতি অনুসারে আউটপুট টোকেন ব্যবহারের বিশদ বিবরণ।
ক্ষেত্র
পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)
টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।
সম্ভাব্য মান
-
textএটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।
-
imageএটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।
-
audioএটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।
-
videoএটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।
-
documentএটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।
মোডালিটির জন্য টোকেনের সংখ্যা।
tool_use_tokens_by_modality অ্যারে (ModalityTokens) (ঐচ্ছিক)
পদ্ধতি অনুসারে টুল-ব্যবহার টোকেন ব্যবহারের একটি বিশদ বিবরণ।
ক্ষেত্র
পদ্ধতি প্রতিক্রিয়া পদ্ধতি (ঐচ্ছিক)
টোকেন সংখ্যার সাথে সংশ্লিষ্ট পদ্ধতি।
সম্ভাব্য মান
-
textএটি নির্দেশ করে যে মডেলটি টেক্সট ফেরত দেবে।
-
imageএটি নির্দেশ করে যে মডেলটি ছবি ফেরত দেবে।
-
audioএটি নির্দেশ করে যে মডেলটি অডিও ফেরত দেবে।
-
videoএটি নির্দেশ করে যে মডেলটি ভিডিও ফেরত দেবে।
-
documentএটি নির্দেশ করে যে মডেলটি ডকুমেন্ট ফেরত দেবে।
মোডালিটির জন্য টোকেনের সংখ্যা।
প্রম্পটের ক্যাশ করা অংশে থাকা টোকেনের সংখ্যা (ক্যাশ করা বিষয়বস্তু)।
প্রম্পটে (প্রসঙ্গ) টোকেনের সংখ্যা।
তৈরি হওয়া সমস্ত প্রতিক্রিয়া জুড়ে টোকেনের মোট সংখ্যা।
চিন্তন মডেলগুলোর জন্য চিন্তার টোকেনের সংখ্যা।
ইন্টারঅ্যাকশন অনুরোধের জন্য মোট টোকেন সংখ্যা (প্রম্পট + প্রতিক্রিয়া + অন্যান্য অভ্যন্তরীণ টোকেন)।
টুল ব্যবহারের নির্দেশনায় উপস্থিত টোকেনের সংখ্যা।
webhook_config WebhookConfig (ঐচ্ছিক)
ঐচ্ছিক। ইন্টারঅ্যাকশন সম্পন্ন হলে নোটিফিকেশন পাওয়ার জন্য ওয়েবহুক কনফিগারেশন।
ক্ষেত্র
ঐচ্ছিক। সেট করা হলে, নিবন্ধিত ওয়েবহুকগুলোর পরিবর্তে এই ওয়েবহুক ইউআরআইগুলো ওয়েবহুক ইভেন্টের জন্য ব্যবহৃত হবে।
ঐচ্ছিক। ব্যবহারকারীর মেটাডেটা যা ওয়েবহুকগুলিতে প্রতিটি ইভেন্ট প্রেরণের সময় ফেরত দেওয়া হবে।
উদাহরণ
উদাহরণ
{ "created": "2025-12-04T15:01:45Z", "id": "v1_ChdXS0l4YWZXTk9xbk0xZThQczhEcmlROBIXV0tJeGFmV05PcW5NMWU4UHM4RHJpUTg", "model": "gemini-3.6-flash", "object": "interaction", "status": "completed", "steps": [ { "type": "model_output", "content": [ { "type": "text", "text": "Hello! I'm doing well, functioning as expected. Thank you for asking! How are you doing today?" } ] } ], "updated": "2025-12-04T15:01:45Z", "usage": { "input_tokens_by_modality": [ { "modality": "text", "tokens": 7 } ], "total_cached_tokens": 0, "total_input_tokens": 7, "total_output_tokens": 23, "total_thought_tokens": 49, "total_tokens": 79, "total_tool_use_tokens": 0 } }
ডেটা মডেল
বিষয়বস্তু
প্রতিক্রিয়ার বিষয়বস্তু।
সম্ভাব্য প্রকার
অডিও কন্টেন্ট
একটি অডিও কন্টেন্ট ব্লক।
অডিও চ্যানেলের সংখ্যা।
অডিও বিষয়বস্তু।
অডিওটির মাইম টাইপ।
সম্ভাব্য মানসমূহ:
-
audio/wavWAV অডিও ফরম্যাট
-
audio/mp3MP3 অডিও ফরম্যাট
-
audio/aiffAIFF অডিও ফরম্যাট
-
audio/aacAAC অডিও ফরম্যাট
-
audio/oggOGG অডিও ফরম্যাট
-
audio/flacFLAC অডিও ফরম্যাট
-
audio/mpegMPEG অডিও ফরম্যাট
-
audio/m4aM4A অডিও ফরম্যাট
-
audio/l16L16 অডিও ফরম্যাট
-
audio/opusOPUS অডিও ফরম্যাট
-
audio/alawALAW অডিও ফরম্যাট
-
audio/mulawMULAW অডিও ফরম্যাট
অডিওটির স্যাম্পল রেট।
কোনো বিবরণ দেওয়া হয়নি।
সর্বদা "audio" তে সেট করা থাকে।
অডিওটির URI।
ডকুমেন্টের বিষয়বস্তু
একটি ডকুমেন্ট কন্টেন্ট ব্লক।
নথির বিষয়বস্তু।
ডকুমেন্টটির মাইম টাইপ।
সম্ভাব্য মানসমূহ:
-
application/pdfপিডিএফ ডকুমেন্ট ফরম্যাট
-
text/csvCSV ডকুমেন্ট ফরম্যাট
কোনো বিবরণ দেওয়া হয়নি।
সর্বদা "document" এ সেট করা থাকে।
ডকুমেন্টটির URI।
ছবির বিষয়বস্তু
একটি চিত্র বিষয়বস্তু ব্লক।
ছবির বিষয়বস্তু।
ছবিটির মাইম টাইপ।
সম্ভাব্য মানসমূহ:
-
image/pngPNG ছবির ফরম্যাট
-
image/jpegJPEG ছবির ফরম্যাট
-
image/webpওয়েবপি ছবির ফরম্যাট
-
image/heicHEIC ছবির ফরম্যাট
-
image/heifHEIF ছবির ফরম্যাট
-
image/gifজিআইএফ ছবির ফরম্যাট
-
image/bmpবিএমপি ছবির ফরম্যাট
-
image/tiffটিআইএফএফ ইমেজ ফরম্যাট
রেজোলিউশন মিডিয়ারেজোলিউশন (ঐচ্ছিক)
গণমাধ্যমের সংকল্প।
সম্ভাব্য মান
-
lowনিম্ন রেজোলিউশন।
-
mediumমাঝারি রেজোলিউশন।
-
highউচ্চ রেজোলিউশন।
-
ultra_highঅতি উচ্চ রেজোলিউশন।
কোনো বিবরণ দেওয়া হয়নি।
সর্বদা "image" হিসেবে সেট করা থাকে।
ছবিটির URI।
টেক্সট কন্টেন্ট
একটি টেক্সট কন্টেন্ট ব্লক।
টীকা অ্যারে (টীকা) (ঐচ্ছিক)
Citation information for model-generated content.
Possible Types
FileCitation
A file citation annotation.
User provided metadata about the retrieved context.
The URI of the file.
End of the attributed segment, exclusive.
The name of the file.
Media ID in-case of image citations, if applicable.
Page number of the cited document, if applicable.
Source attributed for a portion of the text.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
No description provided.
Always set to "file_citation" .
PlaceCitation
A place citation annotation.
End of the attributed segment, exclusive.
Title of the place.
The ID of the place, in `places/{place_id}` format.
review_snippets array (ReviewSnippet) (optional)
Snippets of reviews that are used to generate answers about the features of a given place in Google Maps.
ক্ষেত্র
The ID of the review snippet.
Title of the review.
A link that corresponds to the user review on Google Maps.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
No description provided.
Always set to "place_citation" .
URI reference of the place.
UrlCitation
A URL citation annotation.
End of the attributed segment, exclusive.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
The title of the URL.
No description provided.
Always set to "url_citation" .
The URL.
WordInfo
Word-level ASR annotation for transcription output. Carries the word text, optional timing, and optional speaker attribution.
End of the attributed segment, exclusive.
End offset in time of the word relative to the start of the audio. Present when timestamp_granularities contains "word".
Optional. Speaker label for this word (eg "spk_1", "spk_2"). Present when diarization_mode is set in TranscriptionConfig.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
Start offset in time of the word relative to the start of the audio. Present when timestamp_granularities contains "word".
The transcribed word.
No description provided.
Always set to "word_info" .
Required. The text content.
No description provided.
Always set to "text" .
VideoContent
A video content block.
The video content.
The mime type of the video.
Possible values:
-
video/mp4MP4 video format
-
video/mpegMPEG video format
-
video/mpgMPG video format
-
video/movMOV video format
-
video/aviAVI video format
-
video/x-flvFLV video format
-
video/webmWebM video format
-
video/wmvWMV video format
-
video/3gpp3GPP video format
A user-defined name for this content block. Can be referenced by the model in the final response.
processing MediaProcessing or enum (string) (optional)
How the model processes this video for understanding.
resolution MediaResolution (optional)
The resolution of the media.
Possible values
-
lowLow resolution.
-
mediumMedium resolution.
-
highউচ্চ রেজোলিউশন।
-
ultra_highUltra high resolution.
No description provided.
Always set to "video" .
The URI of the video.
উদাহরণ
অডিও
{ "type": "audio", "data": "BASE64_ENCODED_AUDIO", "mime_type": "audio/wav" }
নথি
{ "type": "document", "data": "BASE64_ENCODED_DOCUMENT", "mime_type": "application/pdf" }
ছবি
{ "type": "image", "data": "BASE64_ENCODED_IMAGE", "mime_type": "image/png" }
পাঠ্য
{ "type": "text", "text": "Hello, how are you?" }
ভিডিও
{ "type": "video", "uri": "https://www.youtube.com/watch?v=9hE5-98ZeCg" }
সরঞ্জাম
A tool that can be used by the model.
Possible Types
CodeExecution
A tool that can be used by the model to execute code.
No description provided.
Always set to "code_execution" .
ComputerUse
A tool that can be used by the model to interact with the computer.
Optional. Disabled safety policies for computer use.
Possible values:
-
financial_transactionsSafety policy for financial transactions.
-
sensitive_data_modificationSafety policy for sensitive data modification.
-
communication_toolSafety policy for communication tools (eg Gmail, Chat, Meet).
-
account_creationSafety policy for account creation.
-
data_modificationSafety policy for data modification.
-
user_consent_managementSafety policy for user consent management.
-
legal_terms_and_agreementsSafety policy for legal terms and agreements.
Whether enable the prompt injection detection check on computer-use request.
The environment being operated.
Possible values:
-
browserOperates in a web browser.
-
mobileOperates in a mobile environment.
-
desktopOperates in a desktop environment.
The list of predefined functions that are excluded from the model call.
No description provided.
Always set to "computer_use" .
FileSearch
A tool that can be used by the model to search files.
The file search store names to search.
Metadata filter to apply to the semantic retrieval documents and chunks.
The number of semantic retrieval chunks to retrieve.
No description provided.
Always set to "file_search" .
Function
A tool that can be used by the model.
A description of the function.
The name of the function.
The JSON Schema for the function's parameters.
No description provided.
Always set to "function" .
GoogleMaps
A tool that can be used by the model to call Google Maps.
Whether to return a widget context token in the tool call result of the response.
The latitude of the user's location.
The longitude of the user's location.
No description provided.
Always set to "google_maps" .
GoogleSearch
A tool that can be used by the model to search Google.
The types of search grounding to enable.
Possible values:
-
web_searchSetting this field enables web search. Only text results are returned.
-
image_searchSetting this field enables image search. Image bytes are returned.
-
enterprise_web_searchSetting this field enables enterprise web search.
No description provided.
Always set to "google_search" .
McpServer
A MCPServer is a server that can be called by the model to perform actions.
allowed_tools array (AllowedTools) (optional)
The allowed tools.
ক্ষেত্র
The mode of the tool choice.
Possible values:
-
autoAuto tool choice.
-
anyAny tool choice.
-
noneNo tool choice.
-
validatedValidated tool choice.
The names of the allowed tools.
Optional: Fields for authentication headers, timeouts, etc., if needed.
The name of the MCPServer.
No description provided.
Always set to "mcp_server" .
The full URL for the MCPServer endpoint. Example: "https://api.example.com/mcp"
Retrieval
A tool that can be used by the model to retrieve files.
exa_ai_search_config ExaAISearchConfig (optional)
Used to specify configuration for ExaAISearch.
ক্ষেত্র
Required. The API key for ExaAiSearch.
Optional. This field can be used to pass any parameter from the Exa.ai Search API.
parallel_ai_search_config ParallelAISearchConfig (optional)
Used to specify configuration for ParallelAISearch.
ক্ষেত্র
Optional. The API key for ParallelAiSearch.
Optional. Custom configs for ParallelAiSearch.
rag_store_config RagStoreConfig (optional)
Used to specify configuration for RagStore.
ক্ষেত্র
rag_resources array (RagResource) (optional)
Optional. The representation of the rag source.
ক্ষেত্র
Optional. RagCorpora resource name.
Optional. rag_file_id. The files should be in the same rag_corpus set in rag_corpus field.
rag_retrieval_config RagRetrievalConfig (optional)
Optional. The retrieval config for the Rag query.
ক্ষেত্র
filter Filter (optional)
Optional. Config for filters.
ক্ষেত্র
Optional. String for metadata filtering.
Optional. Only returns contexts with vector distance smaller than the threshold.
Optional. Only returns contexts with vector similarity larger than the threshold.
hybrid_search HybridSearch (optional)
Optional. Config for Hybrid Search.
ক্ষেত্র
Optional. Alpha value controls the weight between dense and sparse vector search results.
ranking Ranking (optional)
Optional. Config for ranking and reranking.
Optional. The number of contexts to retrieve.
The types of file retrieval to enable.
Possible values:
-
rag_store -
exa_ai_search -
parallel_ai_search
No description provided.
Always set to "retrieval" .
UrlContext
A tool that can be used by the model to fetch URL context.
No description provided.
Always set to "url_context" .
উদাহরণ
CodeExecution
ComputerUse
FileSearch
Function
GoogleMaps
GoogleSearch
McpServer
Retrieval
No examples available for this type.
UrlContext
InteractionSseEvent
Possible Types
Polymorphic discriminator: event_type
ErrorEvent
error Error (optional)
No description provided.
ক্ষেত্র
A URI that identifies the error type.
A human-readable error message.
The event_id token to be used to resume the interaction stream, from this event.
No description provided.
Always set to "error" .
InteractionCompletedEvent
The event_id token to be used to resume the interaction stream, from this event.
No description provided.
Always set to "interaction.completed" .
interaction InteractionSseEventInteraction (required)
Partial completed interaction resource emitted at the end of the stream.
ক্ষেত্র
The agent to interact with.
Output only. The time at which the response was created in ISO 8601 format.
Required. Output only. A unique identifier for the interaction completion.
The model that will complete your prompt.
Output only. The resource type.
service_tier ServiceTier (optional)
The service tier for the interaction.
Possible values
-
flexFlex service tier.
-
standardStandard service tier.
-
priorityPriority service tier.
-
deferredDeferred service tier.
Required. Output only. The status of the interaction.
Possible values:
-
in_progressThe interaction is in progress.
-
requires_actionThe interaction requires action/input from the user.
-
completedThe interaction is completed.
-
failedThe interaction failed.
-
cancelledThe interaction was cancelled.
-
incompleteThe interaction is completed, but contains incomplete results (eg hitting max_tokens).
Output only. The steps that make up the interaction, if included in this event.
Output only. The time at which the response was last updated in ISO 8601 format.
usage Usage (optional)
Output only. Statistics on the interaction request's token usage.
ক্ষেত্র
cached_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of cached token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
grounding_tool_count array (GroundingToolCount) (optional)
Grounding tool count.
ক্ষেত্র
The number of grounding tool counts.
The grounding tool type associated with the count.
Possible values:
-
google_searchGrounding with Google Web Search and Image Search, & Web Grounding for Enterprise.
-
google_mapsGrounding with Google Maps.
-
retrievalGrounding with customer's data, for example, VertexAISearch.
input_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of input token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
output_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of output token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
tool_use_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of tool-use token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
Number of tokens in the cached part of the prompt (the cached content).
Number of tokens in the prompt (context).
Total number of tokens across all the generated responses.
Number of tokens of thoughts for thinking models.
Total token count for the interaction request (prompt + responses + other internal tokens).
Number of tokens present in tool-use prompt(s).
InteractionCreatedEvent
The event_id token to be used to resume the interaction stream, from this event.
No description provided.
Always set to "interaction.created" .
interaction InteractionSseEventInteraction (required)
Partial interaction resource emitted when the stream is created.
ক্ষেত্র
The agent to interact with.
Output only. The time at which the response was created in ISO 8601 format.
Required. Output only. A unique identifier for the interaction completion.
The model that will complete your prompt.
Output only. The resource type.
service_tier ServiceTier (optional)
The service tier for the interaction.
Possible values
-
flexFlex service tier.
-
standardStandard service tier.
-
priorityPriority service tier.
-
deferredDeferred service tier.
Required. Output only. The status of the interaction.
Possible values:
-
in_progressThe interaction is in progress.
-
requires_actionThe interaction requires action/input from the user.
-
completedThe interaction is completed.
-
failedThe interaction failed.
-
cancelledThe interaction was cancelled.
-
incompleteThe interaction is completed, but contains incomplete results (eg hitting max_tokens).
Output only. The steps that make up the interaction, if included in this event.
Output only. The time at which the response was last updated in ISO 8601 format.
usage Usage (optional)
Output only. Statistics on the interaction request's token usage.
ক্ষেত্র
cached_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of cached token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
grounding_tool_count array (GroundingToolCount) (optional)
Grounding tool count.
ক্ষেত্র
The number of grounding tool counts.
The grounding tool type associated with the count.
Possible values:
-
google_searchGrounding with Google Web Search and Image Search, & Web Grounding for Enterprise.
-
google_mapsGrounding with Google Maps.
-
retrievalGrounding with customer's data, for example, VertexAISearch.
input_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of input token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
output_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of output token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
tool_use_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of tool-use token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
Number of tokens in the cached part of the prompt (the cached content).
Number of tokens in the prompt (context).
Total number of tokens across all the generated responses.
Number of tokens of thoughts for thinking models.
Total token count for the interaction request (prompt + responses + other internal tokens).
Number of tokens present in tool-use prompt(s).
InteractionStatusUpdate
The event_id token to be used to resume the interaction stream, from this event.
No description provided.
Always set to "interaction.status_update" .
No description provided.
No description provided.
Possible values:
-
in_progressThe interaction is in progress.
-
requires_actionThe interaction requires action/input from the user.
-
completedThe interaction is completed.
-
failedThe interaction failed.
-
cancelledThe interaction was cancelled.
-
incompleteThe interaction is completed, but contains incomplete results (eg hitting max_tokens).
-
budget_exceededThe interaction was halted because the token budget was exceeded.
-
queuedThe interaction is queued, waiting for processing (eg waiting for off-peak capacity).
StepDelta
delta StepDeltaData (required)
No description provided.
Possible Types
ArgumentsDelta
No description provided.
No description provided.
Always set to "arguments_delta" .
AudioDelta
The number of audio channels.
No description provided.
No description provided.
Possible values:
-
audio/wavWAV অডিও ফরম্যাট
-
audio/mp3MP3 audio format
-
audio/aiffAIFF audio format
-
audio/aacAAC audio format
-
audio/oggOGG audio format
-
audio/flacFLAC audio format
-
audio/mpegMPEG audio format
-
audio/m4aM4A audio format
-
audio/l16L16 audio format
-
audio/opusOPUS audio format
-
audio/alawALAW audio format
-
audio/mulawMULAW audio format
The sample rate of the audio.
No description provided.
Always set to "audio" .
No description provided.
CodeExecutionCallDelta
arguments CodeExecutionCallArguments (required)
No description provided.
ক্ষেত্র
The code to be executed.
Programming language of the `code`.
Possible values:
-
pythonPython >= 3.10, with numpy and simpy available.
A signature hash for backend validation.
No description provided.
Always set to "code_execution_call" .
CodeExecutionResultDelta
No description provided.
No description provided.
A signature hash for backend validation.
No description provided.
Always set to "code_execution_result" .
DocumentDelta
No description provided.
No description provided.
Possible values:
-
application/pdfPDF document format
-
text/csvCSV document format
No description provided.
Always set to "document" .
No description provided.
FileSearchCallDelta
A signature hash for backend validation.
No description provided.
Always set to "file_search_call" .
FileSearchResultDelta
result array (FileSearchResult) (required)
No description provided.
A signature hash for backend validation.
No description provided.
Always set to "file_search_result" .
FunctionResultDelta
Required. ID to match the ID from the function call block.
No description provided.
No description provided.
No description provided.
No description provided.
Always set to "function_result" .
GoogleMapsCallDelta
arguments GoogleMapsCallArguments (optional)
The arguments to pass to the Google Maps tool.
ক্ষেত্র
The queries to be executed.
A signature hash for backend validation.
No description provided.
Always set to "google_maps_call" .
GoogleMapsResultDelta
result array (GoogleMapsResult) (optional)
The results of the Google Maps.
ক্ষেত্র
places array (Places) (optional)
The places that were found.
ক্ষেত্র
Title of the place.
The ID of the place, in `places/{place_id}` format.
review_snippets array (ReviewSnippet) (optional)
Snippets of reviews that are used to generate answers about the features of a given place in Google Maps.
ক্ষেত্র
The ID of the review snippet.
Title of the review.
A link that corresponds to the user review on Google Maps.
URI reference of the place.
Resource name of the Google Maps widget context token.
A signature hash for backend validation.
No description provided.
Always set to "google_maps_result" .
GoogleSearchCallDelta
arguments GoogleSearchCallArguments (required)
No description provided.
ক্ষেত্র
Web search queries for the following-up web search.
A signature hash for backend validation.
No description provided.
Always set to "google_search_call" .
GoogleSearchResultDelta
No description provided.
result array (GoogleSearchResult) (required)
No description provided.
ক্ষেত্র
Web content snippet that can be embedded in a web page or an app webview.
A signature hash for backend validation.
No description provided.
Always set to "google_search_result" .
ImageDelta
No description provided.
No description provided.
Possible values:
-
image/pngPNG image format
-
image/jpegJPEG image format
-
image/webpWebP image format
-
image/heicHEIC image format
-
image/heifHEIF image format
-
image/gifGIF image format
-
image/bmpBMP image format
-
image/tiffTIFF image format
resolution MediaResolution (optional)
The resolution of the media.
Possible values
-
lowLow resolution.
-
mediumMedium resolution.
-
highউচ্চ রেজোলিউশন।
-
ultra_highUltra high resolution.
No description provided.
Always set to "image" .
No description provided.
McpServerToolCallDelta
No description provided.
No description provided.
No description provided.
No description provided.
Always set to "mcp_server_tool_call" .
McpServerToolResultDelta
No description provided.
No description provided.
No description provided.
No description provided.
Always set to "mcp_server_tool_result" .
ProcessingCallDelta
Streaming delta for a server-initiated media processing step.
A signature hash for backend validation.
No description provided.
Always set to "processing_call" .
ProcessingResultDelta
Streaming delta for the result of a server-initiated media processing step.
A signature hash for backend validation.
No description provided.
Always set to "processing_result" .
RetrievalCallDelta
Used by Vertex Retrieval tools such as Parallel AI, Exa AI, Vertex AI Search, etc. RetrievalType decides which tool is used.
arguments RetrievalStepArguments (required)
Required. The arguments to pass to the Retrieval tool.
ক্ষেত্র
Queries for Retrieval information.
The type of retrieval tools.
Possible values:
-
rag_storeThe type of retrieval tools.
-
exa_ai_searchThe type of retrieval tools.
-
parallel_ai_searchThe type of retrieval tools.
A signature hash for backend validation.
No description provided.
Always set to "retrieval_call" .
RetrievalResultDelta
Used by Vertex Retrieval tools such as Parallel AI, Exa AI, Vertex AI Search, etc. ToolResultDelta.type
Whether the retrieval resulted in an error.
A signature hash for backend validation.
No description provided.
Always set to "retrieval_result" .
TextAnnotationDelta
annotations array (Annotation) (optional)
Citation information for model-generated content.
Possible Types
FileCitation
A file citation annotation.
User provided metadata about the retrieved context.
The URI of the file.
End of the attributed segment, exclusive.
The name of the file.
Media ID in-case of image citations, if applicable.
Page number of the cited document, if applicable.
Source attributed for a portion of the text.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
No description provided.
Always set to "file_citation" .
PlaceCitation
A place citation annotation.
End of the attributed segment, exclusive.
Title of the place.
The ID of the place, in `places/{place_id}` format.
review_snippets array (ReviewSnippet) (optional)
Snippets of reviews that are used to generate answers about the features of a given place in Google Maps.
ক্ষেত্র
The ID of the review snippet.
Title of the review.
A link that corresponds to the user review on Google Maps.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
No description provided.
Always set to "place_citation" .
URI reference of the place.
UrlCitation
A URL citation annotation.
End of the attributed segment, exclusive.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
The title of the URL.
No description provided.
Always set to "url_citation" .
The URL.
WordInfo
Word-level ASR annotation for transcription output. Carries the word text, optional timing, and optional speaker attribution.
End of the attributed segment, exclusive.
End offset in time of the word relative to the start of the audio. Present when timestamp_granularities contains "word".
Optional. Speaker label for this word (eg "spk_1", "spk_2"). Present when diarization_mode is set in TranscriptionConfig.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
Start offset in time of the word relative to the start of the audio. Present when timestamp_granularities contains "word".
The transcribed word.
No description provided.
Always set to "word_info" .
No description provided.
Always set to "text_annotation_delta" .
TextDelta
No description provided.
No description provided.
Always set to "text" .
ThoughtSignatureDelta
Signature to match the backend source to be part of the generation.
No description provided.
Always set to "thought_signature" .
ThoughtSummaryDelta
A new summary item to be added to the thought.
No description provided.
Always set to "thought_summary" .
UrlContextCallDelta
arguments UrlContextCallArguments (required)
No description provided.
ক্ষেত্র
The URLs to fetch.
A signature hash for backend validation.
No description provided.
Always set to "url_context_call" .
UrlContextResultDelta
No description provided.
result array (UrlContextResult) (required)
No description provided.
ক্ষেত্র
The status of the URL retrieval.
Possible values:
-
successUrl retrieval is successful.
-
errorUrl retrieval is failed due to error.
-
paywallUrl retrieval is failed because the content is behind paywall.
-
unsafeUrl retrieval is failed because the content is unsafe.
The URL that was fetched.
A signature hash for backend validation.
No description provided.
Always set to "url_context_result" .
VideoDelta
No description provided.
No description provided.
Possible values:
-
video/mp4MP4 video format
-
video/mpegMPEG video format
-
video/mpgMPG video format
-
video/movMOV video format
-
video/aviAVI video format
-
video/x-flvFLV video format
-
video/webmWebM video format
-
video/wmvWMV video format
-
video/3gpp3GPP video format
-
video/jpeg2000JPEG 2000 video format
resolution MediaResolution (optional)
The resolution of the media.
Possible values
-
lowLow resolution.
-
mediumMedium resolution.
-
highউচ্চ রেজোলিউশন।
-
ultra_highUltra high resolution.
No description provided.
Always set to "video" .
No description provided.
The event_id token to be used to resume the interaction stream, from this event.
No description provided.
Always set to "step.delta" .
No description provided.
metadata StepDeltaMetadata (optional)
No description provided.
ক্ষেত্র
total_usage Usage (optional)
Statistics on the interaction request's token usage.
ক্ষেত্র
cached_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of cached token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
grounding_tool_count array (GroundingToolCount) (optional)
Grounding tool count.
ক্ষেত্র
The number of grounding tool counts.
The grounding tool type associated with the count.
Possible values:
-
google_searchGrounding with Google Web Search and Image Search, & Web Grounding for Enterprise.
-
google_mapsGrounding with Google Maps.
-
retrievalGrounding with customer's data, for example, VertexAISearch.
input_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of input token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
output_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of output token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
tool_use_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of tool-use token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
Number of tokens in the cached part of the prompt (the cached content).
Number of tokens in the prompt (context).
Total number of tokens across all the generated responses.
Number of tokens of thoughts for thinking models.
Total token count for the interaction request (prompt + responses + other internal tokens).
Number of tokens present in tool-use prompt(s).
StepStart
The event_id token to be used to resume the interaction stream, from this event.
No description provided.
Always set to "step.start" .
No description provided.
No description provided.
StepStop
The event_id token to be used to resume the interaction stream, from this event.
No description provided.
Always set to "step.stop" .
No description provided.
step_usage Usage (optional)
Model usage stats for this specific step.
ক্ষেত্র
cached_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of cached token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
grounding_tool_count array (GroundingToolCount) (optional)
Grounding tool count.
ক্ষেত্র
The number of grounding tool counts.
The grounding tool type associated with the count.
Possible values:
-
google_searchGrounding with Google Web Search and Image Search, & Web Grounding for Enterprise.
-
google_mapsGrounding with Google Maps.
-
retrievalGrounding with customer's data, for example, VertexAISearch.
input_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of input token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
output_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of output token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
tool_use_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of tool-use token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
Number of tokens in the cached part of the prompt (the cached content).
Number of tokens in the prompt (context).
Total number of tokens across all the generated responses.
Number of tokens of thoughts for thinking models.
Total token count for the interaction request (prompt + responses + other internal tokens).
Number of tokens present in tool-use prompt(s).
usage Usage (optional)
Cumulative model usage stats from the start of the session.
ক্ষেত্র
cached_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of cached token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
grounding_tool_count array (GroundingToolCount) (optional)
Grounding tool count.
ক্ষেত্র
The number of grounding tool counts.
The grounding tool type associated with the count.
Possible values:
-
google_searchGrounding with Google Web Search and Image Search, & Web Grounding for Enterprise.
-
google_mapsGrounding with Google Maps.
-
retrievalGrounding with customer's data, for example, VertexAISearch.
input_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of input token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
output_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of output token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
tool_use_tokens_by_modality array (ModalityTokens) (optional)
A breakdown of tool-use token usage by modality.
ক্ষেত্র
modality ResponseModality (optional)
The modality associated with the token count.
Possible values
-
textIndicates the model should return text.
-
imageIndicates the model should return images.
-
audioIndicates the model should return audio.
-
videoIndicates the model should return video.
-
documentIndicates the model should return documents.
Number of tokens for the modality.
Number of tokens in the cached part of the prompt (the cached content).
Number of tokens in the prompt (context).
Total number of tokens across all the generated responses.
Number of tokens of thoughts for thinking models.
Total token count for the interaction request (prompt + responses + other internal tokens).
Number of tokens present in tool-use prompt(s).
উদাহরণ
Error Event
{ "error": { "code": "not_found", "message": "Failed to get completed interaction: Result not found." }, "event_type": "error" }
Interaction Completed
{ "event_id": "evt_123", "event_type": "interaction.completed", "interaction": { "created": "2025-12-04T15:01:45Z", "id": "v1_ChdXS0l4YWZXTk9xbk0xZThQczhEcmlROBIXV0tJeGFmV05PcW5NMWU4UHM4RHJpUTg", "model": "gemini-3.6-flash", "status": "completed", "updated": "2025-12-04T15:01:45Z" } }
Interaction Completed
{ "event_id": "evt_123", "event_type": "interaction.completed", "interaction": { "created": "2025-12-04T15:01:45Z", "id": "v1_ChdXS0l4YWZXTk9xbk0xZThQczhEcmlROBIXV0tJeGFmV05PcW5NMWU4UHM4RHJpUTg", "model": "gemini-3-flash-preview", "object": "interaction", "status": "completed", "updated": "2025-12-04T15:01:45Z" } }
Interaction Created
{ "event_id": "evt_123", "event_type": "interaction.created", "interaction": { "created": "2025-12-04T15:01:45Z", "id": "v1_ChdXS0l4YWZXTk9xbk0xZThQczhEcmlROBIXV0tJeGFmV05PcW5NMWU4UHM4RHJpUTg", "model": "gemini-3.6-flash", "status": "in_progress", "updated": "2025-12-04T15:01:45Z" } }
Interaction Created
{ "event_id": "evt_123", "event_type": "interaction.created", "interaction": { "id": "v1_ChdXS0l4YWZXTk9xbk0xZThQczhEcmlROBIXV0tJeGFmV05PcW5NMWU4UHM4RHJpUTg", "model": "gemini-3-flash-preview", "object": "interaction", "status": "in_progress" } }
Interaction Status Update
{ "event_type": "interaction.status_update", "interaction_id": "v1_ChdTMjQ0YWJ5TUF1TzcxZThQdjRpcnFRcxIXUzI0NGFieU1BdU83MWU4UHY0aXJxUXM", "status": "in_progress" }
Step Delta
{ "delta": { "type": "text", "text": "Hello" }, "event_type": "step.delta", "index": 0 }
Step Start
{ "event_type": "step.start", "index": 0, "step": { "type": "model_output" } }
Step Stop
{ "event_type": "step.stop", "index": 0 }
ResponseFormat
Possible Types
AudioResponseFormat
Configuration for audio output format.
Bit rate in bits per second (bps). Only applicable for compressed formats (MP3, Opus).
The delivery mode for the audio output.
Possible values:
-
inlineAudio data is returned inline in the response.
-
uriAudio data is returned as a URI.
The MIME type of the audio output.
Possible values:
-
audio/mp3MP3 audio format.
-
audio/ogg_opusOGG Opus audio format.
-
audio/l16Raw PCM (L16) audio format.
-
audio/wavWAV audio format.
-
audio/alawA-law audio format.
-
audio/mulawMu-law audio format.
Sample rate in Hz.
No description provided.
Always set to "audio" .
ImageResponseFormat
Configuration for image output format.
The aspect ratio for the image output.
Possible values:
-
1:11:1 aspect ratio.
-
2:32:3 aspect ratio.
-
3:23:2 aspect ratio.
-
3:43:4 aspect ratio.
-
4:34:3 aspect ratio.
-
4:54:5 aspect ratio.
-
5:45:4 aspect ratio.
-
9:169:16 aspect ratio.
-
16:916:9 aspect ratio.
-
21:921:9 aspect ratio.
-
1:81:8 aspect ratio.
-
8:18:1 aspect ratio.
-
1:41:4 aspect ratio.
-
4:14:1 aspect ratio.
The delivery mode for the image output.
Possible values:
-
inlineImage data is returned inline in the response.
-
uriImage data is returned as a URI.
The size of the image output.
Possible values:
-
512512px image size.
-
1K1K image size.
-
2K2K image size.
-
4K4K image size.
The MIME type of the image output.
Possible values:
-
image/jpegJPEG image format.
No description provided.
Always set to "image" .
TextResponseFormat
Configuration for text output format.
The MIME type of the text output.
Possible values:
-
application/jsonJSON output format.
-
text/plainPlain text output format.
The JSON schema that the output should conform to. Only applicable when mime_type is application/json.
No description provided.
Always set to "text" .
VideoResponseFormat
Configuration for video output format.
The aspect ratio for the video output.
Possible values:
-
16:916:9 aspect ratio.
-
9:169:16 aspect ratio.
The delivery mode for the video output.
Possible values:
-
inlineVideo data is returned inline in the response.
-
uriVideo data is returned as a URI.
The duration for the video output.
The Cloud Storage URI to store the video output. Required for Vertex if delivery mode is URI.
The video output resolution. Defaults to 720p.
Possible values:
-
360p360p resolution.
-
720p720p resolution.
-
1080p1080p resolution.
-
4k4K resolution.
No description provided.
Always set to "video" .
উদাহরণ
Audio Output
{ "type": "audio", "sample_rate": 24000 }
Image Output
{ "type": "image", "aspect_ratio": "16:9", "image_size": "1K", "mime_type": "image/jpeg" }
Text Output (JSON Schema)
{ "type": "text", "mime_type": "application/json", "schema": { "type": "object", "properties": { "ingredients": { "type": "array", "items": { "type": "string" } }, "recipe_name": { "type": "string" } }, "required": [ "ingredients", "recipe_name" ] } }
ভিডিও আউটপুট
{ "type": "video", "aspect_ratio": "16:9", "delivery": "inline" }
ধাপ
A step in the interaction.
Possible Types
CodeExecutionCallStep
Code execution call step.
arguments CodeExecutionCallStepArguments (required)
Required. The arguments to pass to the code execution.
ক্ষেত্র
The code to be executed.
Programming language of the `code`.
Possible values:
-
pythonPython >= 3.10, with numpy and simpy available.
Required. A unique ID for this specific tool call.
A signature hash for backend validation.
No description provided.
Always set to "code_execution_call" .
CodeExecutionResultStep
Code execution result step.
Required. ID to match the ID from the function call block.
Whether the code execution resulted in an error.
Required. The output of the code execution.
A signature hash for backend validation.
No description provided.
Always set to "code_execution_result" .
FileSearchCallStep
File Search call step.
Required. A unique ID for this specific tool call.
A signature hash for backend validation.
No description provided.
Always set to "file_search_call" .
FileSearchResultStep
File Search result step.
Required. ID to match the ID from the function call block.
A signature hash for backend validation.
No description provided.
Always set to "file_search_result" .
FunctionCallStep
A function tool call step.
Required. The arguments to pass to the function.
Required. A unique ID for this specific tool call.
Required. The name of the tool to call.
No description provided.
Always set to "function_call" .
FunctionResultStep
Result of a function tool call.
Required. ID to match the ID from the function call block.
Whether the tool call resulted in an error.
The name of the tool that was called.
Required. The result of the tool call.
No description provided.
Always set to "function_result" .
GoogleMapsCallStep
Google Maps call step.
arguments GoogleMapsCallStepArguments (optional)
The arguments to pass to the Google Maps tool.
ক্ষেত্র
The queries to be executed.
Required. A unique ID for this specific tool call.
A signature hash for backend validation.
No description provided.
Always set to "google_maps_call" .
GoogleMapsResultStep
Google Maps result step.
Required. ID to match the ID from the function call block.
result array (GoogleMapsResultItem) (required)
No description provided.
ক্ষেত্র
places array (GoogleMapsResultPlaces) (optional)
No description provided.
ক্ষেত্র
No description provided.
No description provided.
review_snippets array (ReviewSnippet) (optional)
No description provided.
ক্ষেত্র
The ID of the review snippet.
Title of the review.
A link that corresponds to the user review on Google Maps.
No description provided.
No description provided.
A signature hash for backend validation.
No description provided.
Always set to "google_maps_result" .
GoogleSearchCallStep
Google Search call step.
arguments GoogleSearchCallStepArguments (required)
Required. The arguments to pass to Google Search.
ক্ষেত্র
Web search queries for the following-up web search.
Required. A unique ID for this specific tool call.
The type of search grounding enabled.
Possible values:
-
web_searchSetting this field enables web search. Only text results are returned.
-
image_searchSetting this field enables image search. Image bytes are returned.
-
enterprise_web_searchSetting this field enables enterprise web search.
A signature hash for backend validation.
No description provided.
Always set to "google_search_call" .
GoogleSearchResultStep
Google Search result step.
Required. ID to match the ID from the function call block.
Whether the Google Search resulted in an error.
result array (GoogleSearchResultItem) (required)
Required. The results of the Google Search.
ক্ষেত্র
Web content snippet that can be embedded in a web page or an app webview.
A signature hash for backend validation.
No description provided.
Always set to "google_search_result" .
McpServerToolCallStep
MCPServer tool call step.
Required. The JSON object of arguments for the function.
Required. A unique ID for this specific tool call.
Required. The name of the tool which was called.
Required. The name of the used MCP server.
No description provided.
Always set to "mcp_server_tool_call" .
McpServerToolResultStep
MCPServer tool result step.
Required. ID to match the ID from the function call block.
Name of the tool which is called for this specific tool call.
Required. The output from the MCP server call. Can be simple text or rich content.
The name of the used MCP server.
No description provided.
Always set to "mcp_server_tool_result" .
ModelOutputStep
Output generated by the model.
No description provided.
No description provided.
Always set to "model_output" .
ThoughtStep
A thought step.
A signature hash for backend validation.
summary array (ThoughtSummaryContent) (optional)
A summary of the thought.
Possible Types
ImageContent
An image content block.
The image content.
The mime type of the image.
Possible values:
-
image/pngPNG image format
-
image/jpegJPEG image format
-
image/webpWebP image format
-
image/heicHEIC image format
-
image/heifHEIF image format
-
image/gifGIF image format
-
image/bmpBMP image format
-
image/tiffTIFF image format
resolution MediaResolution (optional)
The resolution of the media.
Possible values
-
lowLow resolution.
-
mediumMedium resolution.
-
highউচ্চ রেজোলিউশন।
-
ultra_highUltra high resolution.
No description provided.
Always set to "image" .
The URI of the image.
TextContent
A text content block.
annotations array (Annotation) (optional)
Citation information for model-generated content.
Possible Types
FileCitation
A file citation annotation.
User provided metadata about the retrieved context.
The URI of the file.
End of the attributed segment, exclusive.
The name of the file.
Media ID in-case of image citations, if applicable.
Page number of the cited document, if applicable.
Source attributed for a portion of the text.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
No description provided.
Always set to "file_citation" .
PlaceCitation
A place citation annotation.
End of the attributed segment, exclusive.
Title of the place.
The ID of the place, in `places/{place_id}` format.
review_snippets array (ReviewSnippet) (optional)
Snippets of reviews that are used to generate answers about the features of a given place in Google Maps.
ক্ষেত্র
The ID of the review snippet.
Title of the review.
A link that corresponds to the user review on Google Maps.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
No description provided.
Always set to "place_citation" .
URI reference of the place.
UrlCitation
A URL citation annotation.
End of the attributed segment, exclusive.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
The title of the URL.
No description provided.
Always set to "url_citation" .
The URL.
WordInfo
Word-level ASR annotation for transcription output. Carries the word text, optional timing, and optional speaker attribution.
End of the attributed segment, exclusive.
End offset in time of the word relative to the start of the audio. Present when timestamp_granularities contains "word".
Optional. Speaker label for this word (eg "spk_1", "spk_2"). Present when diarization_mode is set in TranscriptionConfig.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
Start offset in time of the word relative to the start of the audio. Present when timestamp_granularities contains "word".
The transcribed word.
No description provided.
Always set to "word_info" .
Required. The text content.
No description provided.
Always set to "text" .
No description provided.
Always set to "thought" .
UrlContextCallStep
URL context call step.
arguments UrlContextCallArguments (required)
Required. The arguments to pass to the URL context.
ক্ষেত্র
The URLs to fetch.
Required. A unique ID for this specific tool call.
A signature hash for backend validation.
No description provided.
Always set to "url_context_call" .
UrlContextResultStep
URL context result step.
Required. ID to match the ID from the function call block.
Whether the URL context resulted in an error.
result array (UrlContextResult) (required)
Required. The results of the URL context.
ক্ষেত্র
The status of the URL retrieval.
Possible values:
-
successUrl retrieval is successful.
-
errorUrl retrieval is failed due to error.
-
paywallUrl retrieval is failed because the content is behind paywall.
-
unsafeUrl retrieval is failed because the content is unsafe.
The URL that was fetched.
A signature hash for backend validation.
No description provided.
Always set to "url_context_result" .
UserInputStep
Input provided by the user.
No description provided.
No description provided.
Always set to "user_input" .
উদাহরণ
CodeExecutionCallStep
{ "type": "code_execution_call", "arguments": { "code": "print(sum(range(1, 11)))" }, "id": "code_call_71021" }
CodeExecutionResultStep
{ "type": "code_execution_result", "call_id": "code_call_71021", "result": "55\n" }
FileSearchCallStep
{ "type": "file_search_call", "id": "file_call_88192" }
FileSearchResultStep
{ "type": "file_search_result", "call_id": "file_call_88192" }
FunctionCallStep
{ "name": "get_weather", "type": "function_call", "arguments": { "location": "Boston, MA" }, "id": "call_98231" }
FunctionResultStep
{ "name": "get_weather", "type": "function_result", "call_id": "call_98231", "result": [ { "type": "text", "text": "{\"weather\":\"sunny\"}" } ] }
GoogleMapsCallStep
{ "type": "google_maps_call", "arguments": { "latitude": 37.7749, "longitude": -122.4194 }, "id": "maps_call_39201" }
GoogleMapsResultStep
{ "type": "google_maps_result", "call_id": "maps_call_39201", "result": [ { "name": "Golden Gate Park", "place_id": "ChIJIQBpAG2ahYAR9R7bNdTLg8M", "rating": 4.8 } ] }
GoogleSearchCallStep
{ "type": "google_search_call", "arguments": { "query": "Who won the men's 100m in Paris 2024?" }, "id": "search_call_19201" }
GoogleSearchResultStep
{ "type": "google_search_result", "call_id": "search_call_19201", "result": [ { "title": "Paris 2024 Olympics: Noah Lyles wins men's 100m gold", "url": "https://olympics.com/en/news/paris-2024-noah-lyles-wins-mens-100m-gold", "snippet": "American Noah Lyles won the Olympic men's 100m gold medal in a photo finish." } ] }
McpServerToolCallStep
{ "name": "calculate_tax", "type": "mcp_server_tool_call", "arguments": { "income": 120000, "state": "CA" }, "id": "mcp_call_29012", "server_name": "financial_mcp_server" }
McpServerToolResultStep
{ "type": "mcp_server_tool_result", "call_id": "mcp_call_29012", "result": { "tax_due": 32400 } }
ModelOutputStep
{ "type": "model_output", "content": [ { "type": "text", "text": "The capital of France is Paris." } ] }
ThoughtStep
{ "type": "thought", "signature": "thought_sig_abcd1234", "summary": [ { "type": "text", "text": "The model is searching Google for the capital of France." } ] }
UrlContextCallStep
{ "type": "url_context_call", "arguments": { "urls": [ "https://www.example.com" ] }, "id": "url_call_10219" }
UrlContextResultStep
{ "type": "url_context_result", "call_id": "url_call_10219", "result": [ { "title": "Example Domain", "url": "https://www.example.com", "snippet": "This domain is for use in illustrative examples in documents." } ] }
UserInputStep
{ "type": "user_input", "content": [ { "type": "text", "text": "What is the capital of France?" } ] }
EnvironmentConfig
Configuration for a custom environment.
ক্ষেত্র
Optional. The environment ID for the interaction. If specified, the request will update the existing environment instead of creating a new one.
Network configuration for the environment.
Possible values:
-
disabledTurns all network off.
sources array (Source) (optional)
No description provided.
ক্ষেত্র
The inline content if `type` is `INLINE`.
Optional encoding for inline content (eg `base64`).
The source of the environment. For Cloud Storage, this is the Cloud Storage path. For GitHub, this is the GitHub path.
Where the source should appear in the environment.
No description provided.
Possible values:
-
gcsA Cloud Storage bucket.
-
inlineInline content.
-
repositoryA generic repository. The protocol prefix in the source URL identifies the provider (eg, github://, gcs://).
-
skill_registryA skill resource from the Skill Registry Service. Skill: projects/{project}/locations/{location}/skills/{skill} SkillRevision: projects/{project}/locations/{location}/skills/{skill}/revisions/{revision} Support mounting all skills under a project: projects/{project}/locations/{location}/skills.
No description provided.
Always set to "remote" .
উদাহরণ
Inline Sources
{ "type": "remote", "sources": [ { "type": "inline", "content": "You are a data analyst. Always include visualizations and export results as PDF.", "target": ".agents/AGENTS.md" }, { "type": "inline", "content": "---\nname: slide-maker\ndescription: Create HTML slide decks\n---\n# Slide Maker\n\nWhen asked to create a presentation:\n1. Analyze the input data\n2. Create an HTML slide deck with reveal.js\n3. Save to /workspace/output/slides.html", "target": ".agents/skills/slide-maker/SKILL.md" } ] }
External Sources
{ "type": "remote", "sources": [ { "type": "repository", "source": "https://github.com/my-org/my-skills.git", "target": ".agents/skills" }, { "type": "gcs", "source": "gs://my-bucket/my-folder", "target": "/workspace/data" } ] }
Network Allowlist
{ "type": "remote", "network": { "allowlist": [ { "domain": "pypi.org" }, { "domain": "*.github.com" } ] } }
Proxy Credentials
{ "type": "remote", "network": { "allowlist": [ { "domain": "api.github.com", "transform": { "Authorization": "Bearer YOUR_GITHUB_TOKEN" } } ] } }
EnvironmentNetworkEgressAllowlist
Outbound networking configuration for the sandbox. Accepts an object with an 'allowlist' array to restrict traffic, or the string 'disabled' to turn off all network access. Omit entirely to allow all outbound traffic with no header injection.
Possible Types
object
Outbound networking configuration for the sandbox. When specified, restricts which external domains the sandbox can reach. Omit entirely to allow all outbound traffic with no header injection.
allowlist array (AllowlistEntry) (optional)
List of allowed outbound domains. Only requests to listed domains are permitted. Use [{'domain': '*'}] to allow all domains while still injecting headers on specific ones.
ক্ষেত্র
Domain to allow outbound requests to. Supports wildcards (eg '*.googleapis.com'). Use '*' to allow all domains.
Headers to inject on all outbound requests matching this domain. Accepts a single dict or a list of dicts. The egress proxy injects these automatically.
স্ট্রিং
Turns all network off.
Possible values
-
disabledTurns all network off.
উদাহরণ
উদাহরণ
{ "allowlist": [ { "domain": "github.com", "transform": [ { "Authorization": "Bearer your-token" } ] }, { "domain": "*.googleapis.com" } ] }
ToolChoiceConfig
The tool choice configuration containing allowed tools.
ক্ষেত্র
allowed_tools AllowedTools (optional)
The allowed tools.
ক্ষেত্র
The mode of the tool choice.
Possible values:
-
autoAuto tool choice.
-
anyAny tool choice.
-
noneNo tool choice.
-
validatedValidated tool choice.
The names of the allowed tools.
উদাহরণ
উদাহরণ
{ "allowed_tools": { "mode": "any", "tools": [ "my_tool" ] } }
ImageContent
An image content block.
ক্ষেত্র
The image content.
The mime type of the image.
Possible values:
-
image/pngPNG image format
-
image/jpegJPEG image format
-
image/webpWebP image format
-
image/heicHEIC image format
-
image/heifHEIF image format
-
image/gifGIF image format
-
image/bmpBMP image format
-
image/tiffTIFF image format
resolution MediaResolution (optional)
The resolution of the media.
Possible values
-
lowLow resolution.
-
mediumMedium resolution.
-
highউচ্চ রেজোলিউশন।
-
ultra_highUltra high resolution.
No description provided.
Always set to "image" .
The URI of the image.
উদাহরণ
ছবি
{ "type": "image", "data": "BASE64_ENCODED_IMAGE", "mime_type": "image/png" }
TextContent
A text content block.
ক্ষেত্র
annotations array (Annotation) (optional)
Citation information for model-generated content.
Possible Types
FileCitation
A file citation annotation.
User provided metadata about the retrieved context.
The URI of the file.
End of the attributed segment, exclusive.
The name of the file.
Media ID in-case of image citations, if applicable.
Page number of the cited document, if applicable.
Source attributed for a portion of the text.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
No description provided.
Always set to "file_citation" .
PlaceCitation
A place citation annotation.
End of the attributed segment, exclusive.
Title of the place.
The ID of the place, in `places/{place_id}` format.
review_snippets array (ReviewSnippet) (optional)
Snippets of reviews that are used to generate answers about the features of a given place in Google Maps.
ক্ষেত্র
The ID of the review snippet.
Title of the review.
A link that corresponds to the user review on Google Maps.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
No description provided.
Always set to "place_citation" .
URI reference of the place.
UrlCitation
A URL citation annotation.
End of the attributed segment, exclusive.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
The title of the URL.
No description provided.
Always set to "url_citation" .
The URL.
WordInfo
Word-level ASR annotation for transcription output. Carries the word text, optional timing, and optional speaker attribution.
End of the attributed segment, exclusive.
End offset in time of the word relative to the start of the audio. Present when timestamp_granularities contains "word".
Optional. Speaker label for this word (eg "spk_1", "spk_2"). Present when diarization_mode is set in TranscriptionConfig.
Start of segment of the response that is attributed to this source. Index indicates the start of the segment, measured in bytes.
Start offset in time of the word relative to the start of the audio. Present when timestamp_granularities contains "word".
The transcribed word.
No description provided.
Always set to "word_info" .
Required. The text content.
No description provided.
Always set to "text" .
উদাহরণ
পাঠ্য
{ "type": "text", "text": "Hello, how are you?" }