← 대시보드로
2026-09-04
5개 영상
2026-09-04 06:02 생성
AI News in 5 Mins: GPT-6 Astra
Nate Herk 2026-09-04
AI News in 5 Mins: GPT-6 Astra
↗ 유튜브에서 보기

핵심 요약

OpenAI가 세대적 도약이자 사실상의 AGI 도달로 평가받는 차세대 AI 모델 'GPT-6 Astra'를 공개했습니다. 텍사스 스타게이트 사이트의 GPU 10만 개 이상으로 학습된 역대 최대 규모 모델로, 강력한 성능과 높은 정렬(Alignment) 수준을 자랑합니다. 현재 일부 조직에 우선 제공 중이며, 수일 내 유료 사용자 및 API, AWS를 통해 순차 배포될 예정입니다.

주요 포인트

  • **AGI급 세대적 도약**: OpenAI 사장 그렉 브록먼이 AGI의 도래로 볼 수 있다고 평가한 차세대 모델
  • **역대 최대 규모 학습**: 텍사스 스타게이트 사이트에서 10만 개 이상의 GPU를 투입해 사전학습, 강화학습, 정렬 연구를 집대성
  • **압도적인 벤치마크 및 가성비**: 기존 5.6 Soul 및 Claude Fable 5.1 대비 더 뛰어난 성능과 저렴한 비용 기록
  • **정렬 및 사용자 의도 파악 강화**: 사용자 의도 이해와 모델 행동 제어가 대폭 개선되어 작업 위임 신뢰도 향상
  • **순차 배포 일정**: 데이브레이크 프로그램을 통한 제한적 선공개 후, 수일 내 ChatGPT Plus/Pro/Enterprise 사용자 및 API/AWS로 확대 제공 예정
What do we have right here is GPT6 Astro, which is a new generation of intelligence. And the president, Greg Brockman of OpenAI, called this a generational leap and said could eventually be seen as the arrival of artificial general intelligence or AGI. Now, quick timeline because I think this is pretty funny. September 1st, 1 p.m., we got Fable 5.1, and everyone's been kind of freaking out. It's been all over X and YouTube since it dropped. And then two hours later at 3:30 p.m. on the same day, OpenI made this tweet which said, "As we prepare to release Astra, we're focused on making increasingly capable AI safe and broadly accessible." And they put out this whole blog. It shows some benchmarks about how much better Astra is than 5.6 Soul, which is an amazing model. I've been using codeex with 5.6 Soul to honestly drive most of my day-to-day. Anyways, let's go back to this post. They dropped this video on X today and they only dropped the first like 12 seconds of it. Now, it's really interesting because today, September 3rd, at 10 a.m., pretty much every AI model went down. Claude, OpenAI, Grock, Cursor, and right after these all went down, that's when OpenAI and CHAGBT dropped these two videos on X that obviously blew up because everyone was like, "Oh my gosh, Astra's coming today." And this video right here that they originally dropped was the introduction. It was this first part of this video right here. Create a yellow circle there. Done. Okay, take this and make it look like a rocket. I like this video to take a lot more detail. Your yellow circle is now a window on a rocket. Okay, this is awesome. Now, make a 3D model in opening letter. Let's build a Anyways, if you guys want to watch the rest of that video, you can obviously get to it here at this page. But anyways, what happened? We are introducing GBT6 Astra, the world's most intelligent and aligned model. This model brings together years of research and big bets across pre-training, reinforcement learning, and alignment. Astra was supposedly trained on over a 100,000 GPUs at Texas Stargate site which is apparently the most amount of training that has ever gone into a model and it is rolling out today but only to a limited set of organizations and I think this is through their daybreak program or maybe it's called Daybreak Access but this will be not available to us yet. It said it will be available to us in the coming days to chatb plus pro business and enterprise users and through openi API and AWS. So absolutely cannot wait to get my hands on that. But the benchmarks here are insane how much better this thing is apparently than Claude Fable 5.1 and for cheaper on all of these different benchmarks. It is pretty ridiculous. Astra is our most aligned model with substantial improvements in understanding user intent and model behavior. You can delegate tasks with greater confidence in Astra's judgment. As one way that we test this, we built a new evaluation informed by the hugging face incident that evaluates whether a model facing difficult or impossible task will go beyond its intended scope. And compared to GBG516 Soul, which without production safeguards went beyond the authorized target 48% of the time, Astra didn't do this at all. So it seems to be much more safe, meaning its intent to exploit is much lower. It's also apparently the famous browser use logo from OpenAI, the world's best computer use model, which is insane because GBZ 5.6 Soul is already so so good at browser use and computer use, but apparently it's a new frontier here for computer use. This screen spot pro benchmark, you can see this is insanely high. It obviously is a little bit more expensive than Soul here, but it is scoring like a 92%. And this is basically can the model locate in a screenshot the right areas to click or the right areas that it needs to analyze or look at. And obviously it scores much higher here than Soul as well as Opus 5 and for much cheaper than Opus 5. Now there are so many benchmarks here that are insanely impressive. They're blowing a bunch of other models out of the water here with Astra. But of course that's what every single release blog of a model looks like. It always is the best, the cheapest, the smartest, the fastest, a new frontier, a new step of intelligence. So, I just cannot wait to get my hands dirty and actually put it headto-head against Fable 5.1 and against 5.6 Soul. But, I really think the timeline on all this was really funny, really interesting. These videos dropped and then we saw inside of Codeex, we saw the tags come through GBD6 Astra and GBD6 Astra AON. And then later today is where we saw this quote come out with Greg Brockman and talking about why this matters. And we saw a bunch of different news articles and things. Someone also said that GBT6 somehow has them more hype than GTA 6, which I thought was kind of funny. We see these two blog posts came out and then they actually ended up being taken away. So then if you went 
You Can Now Run Your Entire Email List From Inside Claude!
Nicholas Puru 2026-09-04
You Can Now Run Your Entire Email List From Inside Claude!
↗ 유튜브에서 보기

핵심 요약

옴니센드(Omnisend)가 공식 AI 커넥터를 출시하여 Claude 및 ChatGPT 내부에서 스토어 이메일 리스트를 직접 관리할 수 있게 됨.

별도의 대시보드 접속이나 API 키 설정 없이 자연어 대화만으로 성과 분석, 고객 세그먼트 생성, 캠페인 이메일 작성까지 한 번에 자동 실행 가능함.

주요 포인트

  • **Claude & ChatGPT 연동 지원:** 옴니센드 공식 커넥터를 통해 복잡한 설치나 API 키 없이 약 1분 만에 무료 연결 가능
  • **자연어 기반 작업 수행:** 프롬프트 입력만으로 지난달 최고 수익 캠페인 확인 등 성과 데이터 즉시 분석
  • **자동 세그먼트 생성:** '1회 구매 후 미방문 고객' 등 원하는 조건의 고객군을 대시보드 없이 채팅창에서 바로 추출
  • **이메일 초안 작성 및 발송 준비:** 재구매 유도 캠페인 이메일을 즉시 작성하여 실제 발송 가능한 상태로 옴니센드 내에 반영
  • **단순 리포트 조회를 넘어선 실행:** 분석 보고서 열람에 그치지 않고 실제 이메일 마케팅 실무 작업을 AI가 직접 처리
I probably shouldn't show you this, but you can now run your entire email list from inside of Claude. So, Omnisend just dropped their official AI connector and it plugs your whole store account directly into ChatGPT or Claude. It covers any performance analysis, any segmenting, and even writing the campaigns themselves. So, here's how to set this up. First, just comment "need" and I'll send you the link to get started. It's completely free. And then inside of ChatGPT, open up your connectors, find Omnisend, and just hit connect. And that's the entire setup. It's about a minute. You don't need an API key or anything else to install. And now from there, you can just open up a brand new chat and talk to it. I'm going to say, "Show me the campaign that made me the most money last month, and then build me a segment of everybody who bought once and never came back. And write them an email that brings them back ready to send." And it does all of this inside of Omnisend. It builds the segment, it drafts the email, and you never open the dashboard once. Now, this isn't just reading your report, it's actually doing the work. And if you want this exact setup plus the three prompts that I used, just comment "need" and I'll send it over to you.
Never Pay for AI Again!
Nicholas Puru 2026-09-04
Never Pay for AI Again!
↗ 유튜브에서 보기

핵심 요약

Claude Code와 같은 유료 AI 코딩 도구를 OmniRoute 로컬 라우터와 연동해 완전 무료로 사용하는 방법을 소개한다. 매달 약 15억 개의 무료 토큰을 활용할 수 있으며, 한 번의 설정으로 다양한 코딩 도구에 적용할 수 있다.

주요 포인트

  • **Claude Code 무료 사용**: 비싼 유료 플랜(월 $200 등) 대신 OmniRoute를 통해 무료 모델을 연결해 비용 없이 코딩 작업을 수행할 수 있다.
  • **방대한 무료 토큰 제공**: 무료 공급자들을 연동하여 매월 약 15억(1.5B) 개의 무료 토큰을 활용 가능하다.
  • **로컬 라우터 방식(OmniRoute)**: 로컬 PC에 라우터를 두고 Base URL과 API Key 변경을 지원하는 모든 툴에 범용적으로 연결할 수 있다.
  • **다양한 개발 도구 호환**: Claude Code 외에도 Cursor, Cline, Gemini CLI, DeepSeek 등 Base URL 커스텀을 지원하는 도구에서 동일하게 적용 가능하다.
  • **AI 툴 구조 분리 활용**: 툴의 뼈대(파일 탐색 및 실행 등 앱 레이어)는 그대로 유지하고 두뇌 역할(모델)만 무료 엔드포인트로 교체하는 원리다.
So Cloud Code can now run for completely free where right now as you can see I've got a couple of Claude code sessions running and right here you can see the model that I am using. It isn't Opus or Fable or Sonnet for that matter. This is a free model. It's coming through Omni Route but I'm using it inside of Claude Code. And with this you get around 1 and a half billion free tokens a month and you can use every single one of them without paying anything. So I've been playing around with this. I've been building some things and this is my company's website. I've just rebuilt this completely with this exact setup. Now, Claude Code, as we all know, it's definitely expensive. I'm on the $200 plan and I still run out of credits every single week. So, this is going to be a very practical way around all of that. And on top of that, in this video, I'll show you how exactly to be setting this up. I won't just cover the setup, but I'll also show you the best free providers to be plugging in, how to get all of them for completely free, and then I will be building with it so you can actually see where this holds up, where it doesn't. Now, just one thing before we start because this is going to be changing completely what this is actually worth to you. So, this isn't just a claude code trick. What we're going to be building this is a router that sits on your laptop and anything that lets you type in a base URL and a key can be pointing at. So, claude codeex cursor client Gemini CLI and also the deep sea caress that half of YouTube has been talking about for the last 2 weeks. You'll be able to use this inside of there as well. But I'm just going to be building it once inside of Cloud Code because that's what most of you are already using. That's what I use most often. And then near the end, I'll take that exact same key, I'll show you, drop it into Deep Sea Carness, and that takes about 20 seconds just to show you that you can run this there as well. So, it's one setup and every tool you already have open. And by the way, I'll have a step-by-step guide that will have all the resources and everything that you need to follow along inside of my free school community. So, make sure to check that out if you're interested in getting all of that, plus all of our other free AI resources that we do not publish on our YouTube channel. It'll be inside of there. Now, first things first, go to the link in the description or just search up omniout.online and it'll come to this web page. Okay, so super quick, let me just cover what this actually is and why you should be using it and how it's different from anything else. Now, up to this point, every AI coding tool is two separate things stacked together. So, there's typically the model, which is the brain. That's the part that is doing all of the thinking, all the strategizing, and then there's the application wrapped around it, which does all the actual work. So this piece it's going to be reading your files, running any sort of commands, editing your code, remembering what you actually had asked for. Now clawed code, this is the application. This is the harness. Opus, you should know that's just the brain. So they ship inside of one box. So most people, they assume they're just one thing altogether. However, on the flip side, they're actually two different things and they're only loosely attached where clawed code it sends every request to an address and that address is just a setting inside of a file. So, if you change the address, everything about clawed code, it stays exactly the same. So, you'll have the same terminal, you'll have the same commands, the same file editing and the thinking it starts coming from just somewhere else completely different. So, that's the entire trick with this. You keep the body, however, you're just swapping the brain here. So with that, where do the free brains actually come from? Now, every one of these companies, we have Open Router. I've covered this several times in the past. We also have Nvidia and about 50 more. So, doesn't matter. These hand out a chunk of free API usage literally every single day. They do it because they want developers building on their platform. They want people coming to them. So, this is just a published free tier with a page on their website. And this isn't a loophole that somebody just had found. However, the catch with this is that any one of them on its own is nothing. So, you are legitimately going to be burning through a day's worth in an afternoon and you're sitting there waiting and waiting. So, Omni Route, this is the piece that actually fixes that. So, you don't have to, you know, do all this configuring and go to another one and, you know, just deal with all the headache. This, however, it's just a small app that runs on your laptop and it holds all of those free keys inside of one place and when a request actually comes in, it hands it to a provider that still has some room left. Now, when that one actually d
I Had Fable 5.1 and 5 Build Me the Same App
Nate Herk 2026-09-04
I Had Fable 5.1 and 5 Build Me the Same App
↗ 유튜브에서 보기

핵심 요약

Fable 5.1과 Fable 5에게 동일한 프롬프트를 제공해 장애 대응 시뮬레이션 앱(OpsFlow)을 제작하게 한 실험 결과다. 두 에이전트는 프롬프트를 완전히 다르게 해석하여 개발 비용($1,200 vs $500)과 작업 시간(1.5일 vs 반나절), UI 및 완성도에서 큰 차이를 보였다.

주요 포인트

  • **동일한 위임형 프롬프트 부여**: 에이전트가 직접 코딩하지 않고, 전략·기획을 총괄하며 하위 워커(Opus: 설계/리뷰, Sonnet: 구현/테스트)에 작업을 위임하도록 지시함
  • **비용 및 작업 시간의 극명한 차이**: 한 모델은 $1,200의 비용과 1.5일의 시간이 걸린 반면, 다른 모델은 $500의 비용으로 반나절 만에 작업을 완료함
  • **제작 대상 앱(OpsFlow)**: n8n 같은 실행 자동화 툴이 아닌, 장애 대응 시나리오를 플로우차트로 그리고 시각적으로 리허설할 수 있는 로컬 퍼스트 시뮬레이션 스튜디오
  • **완성 기준 엄격 적용**: 그럴듯한 구현에 그치지 않고 완벽히 동작하며 발표 가능한 수준의 객관적 증거가 나올 때까지 오케스트레이션을 지속하도록 요구함
So, I just had Fable 5.1 and Fable 5 build me the exact same app, and it wasn't even close. Not only did these apps look and feel very different, but one cost me 1,200 bucks, and one cost me 500 bucks. One worked for a day and a half, one worked for just half a day. So, today I'm going to be breaking all of this down for you guys. So, let's not waste any time and just get straight into the video. All right, so I thought that I would start with the prompt that I gave these two agents. Now, this was a big/goal prompt. Please don't roast my prompting. But, I gave them both the exact same prompt. The only difference is that this one said you are Fable 5.1 and you work inside of this folder whereas the other one said you are Fable 5 and you work inside of this folder. That's the only difference. Everything else word for word the exact same. And I'm going to show you guys how interesting it is because they interpreted it so differently. But here's something that I think is key. I didn't have Fable 5.1 build the entire app and Fable 5 build their entire app. I wanted them to basically delegate everything. I said you own strategy, planning, delegation, sequencing, quality standards, and final acceptance. Do not personally perform engineering research, debugging, testing, or visual design production. Preserve your context by delegating execution through dynamic workflows. Primarily use opus workers for architecture, product and design direction, difficult problem solving, and independent reviews. Primarily use sonnet workers for implementation, research, testing, debugging, and iteration. Give workers clear ownership, prevent conflicting edits, and replace or redirect workers when results are weak. Worker failure is still your responsibility. So, I'm not going to read this entire prompt, but this is kind of the gist of it. I gave it bullets on the required product and I gave it the definition of done and I told it to not stop at a plausible implementation but continue orchestrating until you have objective evidence that the complete product works and is presentation ready. So essentially what these agents are building is they are building ops flow which is a local first visual automation studio for designing and simulating incident response workflows. So it's not an workflow automation platform like nen. It is just meant to visually show and kind of like rehearse. It's for the person who owns what happens when production breaks. So, in layman terms, it lets you draw your emergency plan as a flowchart and then you can actually visually watch how that plays out. So, now that that boring stuff is out of the way, let's take a look at the actual apps that they built. Okay, so here is the first one. I'm not going to tell you who made this one yet. I want you guys to try to form your own opinions and then I'll reveal that after we look at both of them. But this is what we're looking at. This is the UI. The first thing I'm thinking about is, is this overwhelming? I am just starting to look at this. I've never ran either of these, but I'm just taking a look now. And let's see what we've got. So, we've got triggers on the lefth hand side. We have trigger, condition, action, approval, resolution. We can drag them into the canvas here, and we can hopefully start to connect them into things. Okay, that looks seems like it works good. We can connect the trigger into this section over here. Okay, so so far the UI is responsive and I can use this. Let's see. Can I close this? Okay, there we go. So this is kind of nodebased workflow but it is vertical rather than horizontal. So what I'm going to do is see if I can delete that connection. Delete this connection. Delete that. Okay. So let's go ahead and run this. So I'm just going to click run right here. We can visually see the stuff happening. That's pretty cool. We're seeing this stuff being processed. And now what do we get at the end? We have a completed run and we're basically able to see what happened. Can I move this up? So I'm I'm not able to expand this. I think that what you should be able to do is grab here and drag this up so you could look at this full screen because this is a little bit hard to actually look at right here. So there was three that got skipped in this branch and there was seven that got executed. The resolution was mitigated by roll back. Okay, run has been completed. Okay, awesome. So that's kind of how this one flowed. This seems to work just fine. We cannot zoom in on that. If we want to zoom in on this thing, oops, sorry guys. We have to go like this, but then we can navigate from there. Okay, so far so good. Let's go into the other version now. So, this is what the other one looks like. This is the UI. If you think it's better, if you think it's worse, I personally think that this UI is a little bit better because I don't know if you guys have realized this sort of card where you've got like the rounded crop color that is just that that screams AI ge
GPT-6 Astra Just Broke the Economics of AI
Nicholas Puru 2026-09-04 설명 기반
GPT-6 Astra Just Broke the Economics of AI
↗ 유튜브에서 보기

핵심 요약

OpenAI의 차세대 모델 GPT-6 Astra가 압도적인 성능 향상과 극단적인 비용 절감을 동시에 달성하며 기존 AI 생태계의 단위 경제학(Unit Economics)을 완전히 뒤흔들고 있다. 단순 프롬프트 기반 질의응답을 넘어 자율적 컴퓨터 제어와 다단계 워크플로우 실행이 가능해지면서 기존 AI 서비스 및 소프트웨어 시장의 가격 구조가 재편되고 있다.

주요 포인트

  • **파괴적인 비용 효율성**: 이전 세대 및 경쟁 모델 대비 성능 대비 API/연산 단가가 대폭 낮아져 AI 도입의 한계 비용이 급감함
  • **자율 컴퓨터 제어(Computer Use) 고도화**: 사람이 직접 지시하던 단계를 넘어 시스템 조작, 코딩, 데이터 분석 등 복잡한 실무를 끝까지 자율 완수함
  • **AI 래퍼(Wrapper) 비즈니스 위협**: 단순 UI 래퍼나 중간 브릿지 서비스의 효용을 없애고 엔드투엔드 자동화 비용을 파괴함
  • **비즈니스 생산성 및 인건비 구조 재편**: 전문 직무 전반에서 에이전트 기반 업무 자동화가 가속되며 기업의 운영 마진과 AI 활용 전략의 근본적 전환을 요구함