Загрузка видео...

Не удалось загрузить видео

На главную

Microsoft has released a major update to Copilot You can now use o3-mini without limit by clicking on “Think Deeper” People sleep on it but Copilot has (for free): → Unlimited reasoning model → Unlimited voice mode → Real-time data w/ GPT-4o → Image gen (Dall-E 3 though)

55,877 просмотров • 1 год назад •via X (Twitter)

Комментарии: 12

Фото профиля Paul Couvert
Paul Couvert1 год назад

- Just access Copilot (web or mobile app) - Select "Think Deeper" in the text field - Copilot will use o3-mini to answer you Once again, no limits. →

Фото профиля ARK Electronics
ARK Electronics2 лет назад

Elevate your drone game with our USA-made, NDAA-compliant flight controllers! Trust in reliable technology that not only enhances your flights but also supports US drone manufacturing capability. 🇺🇸✈️ #USAMade #NDAA #drones #uav #uas #px4 #ardupilot #USA #unmanned #opensource

Фото профиля Sai Rahul
Sai Rahul1 год назад

Ah yes. I completely forgot the copilot 😅

Фото профиля Paul Couvert
Paul Couvert1 год назад

Good thing I'm here to remind you of its existence then 😂

Фото профиля Josh Marino
Josh Marino1 год назад

So people are paying $200 a month for unlimited voice mode but they could get it with Co-Pilot for free?

Фото профиля Paul Couvert
Paul Couvert1 год назад

I believe this is also the case on ChatGPT Plus but I don't know if it's unlimited or if they just increased the limit.

Фото профиля Shushant Lakhyani
Shushant Lakhyani1 год назад

There's no need of ChatGPT's subscription now

Фото профиля Paul Couvert
Paul Couvert1 год назад

Depending on the task, but it can replace it in many situations!

Фото профиля MadMonke.sol
MadMonke.sol1 год назад

accessing copilot opens a world of creativity, doesn’t it? excited to see the insights we'll uncover together.

Фото профиля Paul Couvert
Paul Couvert1 год назад

Worth a try!

Фото профиля Prometheus
Prometheus1 год назад

Is this a desktop app? Where to download it?

Фото профиля Paul Couvert
Paul Couvert1 год назад

More like a progressive web app but a native one is available in preview and should be available soon.

Похожие видео

Cerebras inference is very fast. So fast that it changes how we think about configuring our LLMs for voice agent use cases. Kimi K2.6 is a 1T parameter reasoning model that Cerebras serves at 650 - 1,000 tokens per second (end-to-end throughput), with time to first token metrics as low as 150ms (latency). These numbers are two to three times faster than other similarly capable models. The biggest lever we get from this kind of speed is that we can use the model in reasoning mode, and still have excellent "time to first non-thinking token." This solves a big pain point we have in 2026 for voice agent use cases. Almost all recent innovation in post-training has focused on making models good at reasoning ("test time compute"). This is great, but it makes the user-facing model latency much, much slower. Which is a problem for conversational voice agents. We can run Kimi K2.6 with reasoning turned on, and get responses faster than other models produce with reasoning disabled. On my 30-turn voice agent benchmark, Kimi K2.6 with reasoning enabled ties GPT 5.1 and Haiku 4.5 with reasoning disabled, and is still about 200ms seconds faster! On my primary task agent benchmark, Kimi K2.6 is now the #2 model. It ranks just behind Gemini 3.5 Flash in "high" reasoning mode, and tied with GLM 5, Sonnet 4.6, and GPT 5.4 with reasoning set to "low." But Kimi K2.6 completes each turn in the agent loop in under 500ms. The other four models are all at least 3x slower. (Models only qualify for this benchmark if they can complete task turns at a P50 <4s.) A couple of other things that this speed buys us, for production voice agents: - Tool calls happen fast enough that we don't have to work around tool call latency in our pipeline design. - We can prompt the model to output structured data at the beginning of a response, followed by plain text for voice generation. This opens up possibilities like asking the model to do complex classification/generation tasks that influence the rest of the pipeline. For example, the model could create a detailed style prompt for a steerable TTS model, for each individual conversation turn. And, of course, you can use Kimi K2.6 with reasoning turned off. Cerebras calls this "instant" mode. Here's a video of a Cerebras Kimi K2.6 voice agent with voice-to-voice response time, measured at the client, under 500ms. This is the true response latency as perceived by the user, including all network and audio codec overhead, transcription and turn detection, Kimi K2.6 token generation, and voice generation. 500ms is, effectively, instant. So the Cerebras naming for this mode is a propos. :-)

kwindla

40,593 просмотров • 4 месяцев назад