컨텐츠로 건너뛰기

Seshat

  • Insights
  • News
  • Knowledge
    • Law & Practice
    • BNF
  • Topics
    • AI & Technology
    • Ideas & Society
  • About

seshat

오래 일할 수 있는 AI와 믿고 맡길 수 있는 AI는 다르다 — Stanford CS329A Part 8

9월 29, 2026 작성자: seshat

Stanford CS329A Part 8을 따라 METR, GDPval, DeepScholar-Bench를 비교하며 긴 task horizon과 실제로 믿고 맡길 수 있는 AI의 차이를 설명한다.

카테고리 AI·테크, Insights 태그 AI 에이전트, Insight Publishing, 미래 기술, 인공지능

Long-Horizon AI Is Not the Same as Reliable AI — Stanford CS329A Part 8

9월 29, 2026 작성자: seshat

Stanford CS329A Part 8 examines long-horizon agent evaluation through METR, GDPval, and DeepScholar-Bench, arguing that longer task horizons are not the same as reliable, delegable AI.

카테고리 AI & Technology, Insights 태그 AI Agents, Artificial Intelligence, Future of Technology, Insight Publishing

탐색만 늘린다고 AI가 똑똑해지는 것은 아니다 — Stanford CS329A Part 7

9월 29, 20269월 29, 2026 작성자: seshat

Stanford CS329A Part 7을 따라 AlphaCode, AlphaCode 2, Search-o1, Search-R1을 연결한다. Search의 핵심은 더 많이 찾는 것이 아니라 더 잘 생성·선별·압축하는 데 있다.

카테고리 AI·테크, Insights 태그 AI 에이전트, Insight Publishing, 미래 기술, 인공지능

Search Is Not Enough: How AI Agents Learn to Find the Right Answer — Stanford CS329A Part 7

9월 29, 20269월 29, 2026 작성자: seshat

Stanford CS329A Part 7 connects AlphaCode, AlphaCode 2, Search-o1, and Search-R1 to show why search only helps when generation, selection, and context refinement improve together.

카테고리 AI & Technology, Insights 태그 AI Agents, Artificial Intelligence, Future of Technology, Insight Publishing

AI는 자기 추론에서 어떻게 다시 학습하는가 — Stanford CS329A Part 6 Train-Time Scaling

9월 29, 20269월 29, 2026 작성자: seshat

Stanford CS329A Part 6를 따라 STaR, DeepSeekMath/GRPO, DAPO를 연결한다. 검증된 reasoning이 어떻게 지속적인 train-time improvement로 바뀌는지 살펴본다.

카테고리 AI·테크, Insights 태그 AI 에이전트, Insight Publishing, 미래 기술, 인공지능

Train-Time Scaling: How AI Learns from Its Own Reasoning — Stanford CS329A Part 6

9월 29, 20269월 29, 2026 작성자: seshat

Stanford CS329A Part 6 connects STaR, DeepSeekMath/GRPO, and DAPO to show how verified reasoning can be turned into persistent train-time improvement.

카테고리 AI & Technology, Insights 태그 AI Agents, Artificial Intelligence, Future of Technology, Insight Publishing

AI 에이전트의 능력과 신뢰성은 다르다 — Stanford CS329A Part 5 Agent Evaluation

9월 29, 20269월 29, 2026 작성자: seshat

Stanford CS329A Part 5의 METR, GDPval, DeepScholarBench를 통해 agent capability와 reliability, context, 전문가 수준 산출물 품질이 왜 별개인지 살펴본다.

카테고리 AI·테크, Insights 태그 AI 에이전트, Insight Publishing, 미래 기술, 인공지능

Capability Is Not Reliability — Stanford CS329A Part 5 Agent Evaluation

9월 29, 20269월 29, 2026 작성자: seshat

Stanford CS329A Part 5 compares METR, GDPval, and DeepScholarBench to show why agent capability, reliability, context, and real-world deliverable quality must be evaluated separately.

카테고리 AI & Technology, Insights 태그 AI Agents, Artificial Intelligence, Future of Technology, Insight Publishing

AI 에이전트의 피드백은 어디서 와야 하는가 — Stanford CS329A Part 4

9월 29, 20269월 29, 2026 작성자: seshat

Stanford CS329A Part 4를 따라 ReAct, 코드 실행 피드백, Constitutional AI를 하나의 질문으로 연결한다. Self-improving agent는 어떤 교정 신호를 믿어야 하는가?

카테고리 AI·테크, Insights 태그 AI 에이전트, Insight Publishing, 미래 기술, 인공지능

Where Should an AI Agent Get Its Feedback? — Stanford CS329A Part 4

9월 29, 20269월 29, 2026 작성자: seshat

Stanford CS329A Part 4 connects ReAct, execution feedback, and Constitutional AI to one question: where should a self-improving agent get corrective signals it can actually trust?

카테고리 AI & Technology, Insights 태그 AI Agents, Artificial Intelligence, Future of Technology, Insight Publishing
이전 글
새 글
← 이전 페이지1 … 페이지3 페이지4 페이지5 페이지6 다음 →

Recent Posts

  • 건강한 노화 10년의 반환점, 이제 문제는 정책이 아니라 실행이다
  • Healthy Ageing at the Midpoint: The Hard Part Is Turning Policy Into Care
  • AI 악용의 핵심 변화는 ‘완전 자율화’보다 비용 하락일 수 있다
  • AI Misuse Is Becoming an Economics Problem Before It Becomes an Autonomy Problem
  • 에이전틱 AI 인프라 경쟁은 ‘칩’보다 ‘운영 루프’로 이동한다

Recent Comments

보여줄 댓글이 없습니다.
© 2026 Seshat • 제작됨 GeneratePress