GPT-Live 1
OpenAI's full-duplex voice frontend for natural conversation that delegates deeper reasoning and actions to a separate backend agent.
OpenAI's full-duplex voice frontend for natural conversation that delegates deeper reasoning and actions to a separate backend agent.
نظرة عامة على اللغة الإنجليزية البسيطة
GPT-Live 1 listens and speaks at the same time, so it can handle pauses, acknowledgements, interruptions, and background noise without forcing every exchange into rigid turns. It also provides transcripts and response text alongside the live audio.
The architecture separates conversation from deeper work. GPT-Live 1 manages the voice experience while a Responses model or an application-operated agent handles reasoning and approved tools. The published voice price therefore excludes backend model usage, tools, telephony, and application infrastructure.
مقارنة الفئة
هذه مواصفات منشورة من قبل الموفر، وليست درجات معيار Cody. اتبع المصادر المرتبطة للحدود الحالية والاستثناءات الخاصة بنقطة النهاية.
الوصول إلى واجهة برمجة التطبيقات والموفر
التوفر
Available through OpenAI's Live API for browser, server, and telephony voice applications.
Direct access follows OpenAI's live supported-country list; the Free API tier is not supported for this model.
تحقق من التوفر المباشرالبيانات والتدريب
يقول OpenAI أن محتوى واجهة برمجة التطبيقات (API) لا يُستخدم للتدريب بشكل افتراضي. قد يتم الاحتفاظ بسجلات مراقبة إساءة الاستخدام الافتراضية لمدة تصل إلى 30 يومًا، مع توفر عناصر تحكم إضافية للمؤسسات المؤهلة.
هذه قراءة موجزة لمواد المزود المذكورة، وليست نصيحة قانونية. يمكن أن تحتوي بوابة الطرف الثالث على شروط تخزين وتوجيه وتدريب وإقامة مختلفة من واجهة برمجة التطبيقات المباشرة الخاصة بصانع النموذج.
اقرأ سياسة المزودالأسئلة المتداولة
OpenAI's full-duplex voice frontend for natural conversation that delegates deeper reasoning and actions to a separate backend agent. GPT-Live 1 listens and speaks at the same time, so it can handle pauses, acknowledgements, interruptions, and background noise without forcing every exchange into rigid turns. It also provides transcripts and response text alongside the live audio.
تم إصدار GPT-Live 1 على 10 سبتمبر 2026 وفقًا لمواد الموفر المذكورة.
Available through OpenAI's Live API for browser, server, and telephony voice applications. مسارات الوصول المدرجة في هذا الدليل هي OpenAI.
$0.05 / voice minute. Voice sessions are billed per second without whole-minute rounding. Backend model and tool usage is billed separately, as are telephony and application infrastructure.
Available through OpenAI's Live API for browser, server, and telephony voice applications. Direct access follows OpenAI's live supported-country list; the Free API tier is not supported for this model.
يقول OpenAI أن محتوى واجهة برمجة التطبيقات (API) لا يُستخدم للتدريب بشكل افتراضي. قد يتم الاحتفاظ بسجلات مراقبة إساءة الاستخدام الافتراضية لمدة تصل إلى 30 يومًا، مع توفر عناصر تحكم إضافية للمؤسسات المؤهلة. تنتمي السياسة إلى مسار الموفر وشروط الحساب، لذا تحقق منها مرة أخرى قبل استخدام الإنتاج.
المقارنات ذات الصلة
نموذج OpenAI للاستدلال المنطقي لتحويل الكلام إلى كلام في الوقت الفعلي لوكلاء الصوت الذين يستخدمون الأدوات والذين يحتاجون أيضًا إلى سياق النص والصورة.
فتح المقارنة الكاملةنموذج معاينة الصوت إلى الصوت الخاص بـ Google للحوار منخفض زمن الوصول مع الوعي متعدد الوسائط والتفكير وتأريض البحث واستدعاء الوظائف.
فتح المقارنة الكاملةنموذج ElevenLabs المعبّر لتحويل النص إلى كلام في الوقت الفعلي للحوار الطبيعي والتوصيل العاطفي والعلامات الصوتية وأكثر من 70 لغة.
فتح المقارنة الكاملةأحدث نموذج تعبيري لتحويل النص إلى كلام في العالم في الوقت الفعلي، مصمم لتذكر تسليم المحادثة والتحدث بأكثر من 100 لغة.
فتح المقارنة الكاملةنموذج SpaceXAI الحالي لتحويل الكلام إلى كلام في الوقت الفعلي لوكلاء المحادثة دون الثانية مع إمكانية الوصول إلى الأداة.
فتح المقارنة الكاملة