<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0">
  <channel>
    <title>Tae-Hoon Yong - AI Engineer</title>
    <link>https://th-yong.com</link>
    <description>Tae-Hoon Yong - AI Engineer</description>
    <language>ko</language>
    <item>
      <title>LiteLLM(Python) 대신 Go로 SSE 중계를 고민할 때 먼저 봐야 할 차이</title>
      <link>https://th-yong.com/tech/litellm-python-vs-go-sse</link>
      <guid>https://th-yong.com/tech/litellm-python-vs-go-sse</guid>
      <pubDate>Wed, 29 Jul 2026 00:00:00 GMT</pubDate>
      <description>LLM 스트리밍 서버가 버거워지는 이유는 연결 수보다 연결을 오래 붙잡는 방식에 있습니다 LLM API를 프록시처럼 중계할 때 가장 먼저 부딪히는 질문은 이것입니다. “응답 생성은 느린데, 왜 서버 자원이 이렇게 빨리 차지?” 핵심 원인은 SSE(Se</description>
    </item>
    <item>
      <title>Memory OS of AI Agent: AI 에이전트 기억을 운영체제처럼 관리하는 3단 메모리 프레임워크</title>
      <link>https://th-yong.com/tech/memory-os-of-ai-agent</link>
      <guid>https://th-yong.com/tech/memory-os-of-ai-agent</guid>
      <pubDate>Tue, 14 Jul 2026 00:00:00 GMT</pubDate>
      <description>1. 개요 Memory OS of AI Agent는 AI 에이전트의 기억을 단기 기억(STM, ShortTerm Memory) · 중기 기억(MTM, MidTerm Memory) · 장기 개인 기억(LPM, LongTerm Personal Memory</description>
    </item>
    <item>
      <title>ORCA: cmux 대안으로 주목받는 고성능 AI 네이티브 터미널</title>
      <link>https://th-yong.com/tech/orca-ai-native-terminal</link>
      <guid>https://th-yong.com/tech/orca-ai-native-terminal</guid>
      <pubDate>Mon, 13 Jul 2026 00:00:00 GMT</pubDate>
      <description>1. 개요 ORCA는 기존 터미널 멀티플렉서(multiplexer) 사용 경험에서 자주 지적되던 느린 반응성, 터미널 표시 오류, 복잡한 세션 관리 문제를 개선하면서, AI 워크플로까지 자연스럽게 통합한 차세대 터미널 도구다. ORCA는 Stabili</description>
    </item>
    <item>
      <title>AI 코드 감사는 어디까지 유효한가: Cloudflare CIRCL 사례가 보여준 현실</title>
      <link>https://th-yong.com/tech/ai-code-audit-circl-reality</link>
      <guid>https://th-yong.com/tech/ai-code-audit-circl-reality</guid>
      <pubDate>Wed, 08 Jul 2026 00:00:00 GMT</pubDate>
      <description>요즘 AI가 코드를 읽고, 버그를 찾고, 심지어 보안 취약점까지 발견할 수 있다는 이야기는 더 이상 낯설지 않다. 하지만 막상 현업의 시선으로 보면 질문은 조금 더 현실적이다. 정말 믿을 만한가, 그리고 어디까지 맡길 수 있는가. 이 질문에 꽤 구체적</description>
    </item>
    <item>
      <title>Escalate hard decisions with the advisor tool: 비싼 추론은 ‘결정 순간’에만 올리는 Claude Code 비용 최적화 패턴</title>
      <link>https://th-yong.com/tech/escalate-hard-decisions-with-advisor-tool</link>
      <guid>https://th-yong.com/tech/escalate-hard-decisions-with-advisor-tool</guid>
      <pubDate>Tue, 07 Jul 2026 00:00:00 GMT</pubDate>
      <description>1. 개요 Escalate hard decisions with the advisor tool은 Claude Code에서 항상 강한 모델을 돌리는 대신, “계획 확정/반복 에러/완료 판정”처럼 결정이 어려운 순간에만 더 강한 모델(Advisor)을 호출</description>
    </item>
    <item>
      <title>스테이블코인·프로그래머블 머니·AI Agent 결제가 여는 온체인 금융 인프라</title>
      <link>https://th-yong.com/tech/stablecoin-programmable-money-ai-agent-onchain-finance</link>
      <guid>https://th-yong.com/tech/stablecoin-programmable-money-ai-agent-onchain-finance</guid>
      <pubDate>Thu, 02 Jul 2026 00:00:00 GMT</pubDate>
      <description>요즘 디지털자산을 둘러싼 이야기를 들으면, 예전과는 결이 꽤 달라졌다는 생각이 든다. 한동안 시장의 관심은 어떤 코인이 오를지, 비트코인이 어디까지 갈지, 제도권 편입이 얼마나 빨라질지 같은 질문에 쏠려 있었다. 그런데 최근에는 조금 더 구조적인 이야</description>
    </item>
    <item>
      <title>Cmux: Agent 시대를 위한 CLI 멀티플렉서 + AI 터미널 워크플로우</title>
      <link>https://th-yong.com/tech/cmux-agent-terminal-multiplexer</link>
      <guid>https://th-yong.com/tech/cmux-agent-terminal-multiplexer</guid>
      <pubDate>Mon, 20 Apr 2026 00:00:00 GMT</pubDate>
      <description>이제 iterms에서 넘어갈 때가 된 것 같습니다. ghostty, warp 가 유행일 때에도 iterms가 익숙한 나머지 바꿀 생각을 안했는데(1주일정도 써보다가 돌아옴) cmux는 Agent시대에 안쓸수가 없는듯 싶습니다. 1. 개요 Cmux는 터</description>
    </item>
    <item>
      <title>Rust, Python, TypeScript: AI 시대의 새로운 프로그래밍 3대장(Trifecta) 보고 느낀 점</title>
      <link>https://th-yong.com/tech/rust-python-typescript-trifecta</link>
      <guid>https://th-yong.com/tech/rust-python-typescript-trifecta</guid>
      <pubDate>Fri, 03 Apr 2026 00:00:00 GMT</pubDate>
      <description>1. 개요 Rust, Python, TypeScript가 “새로운 프로그래밍 3대장(Trifecta)”로 묶이며, 앞으로 소프트웨어 개발의 중심 언어로 부상할 가능성이 커지고 있습니다. 특히 LLM(Large Language Model) 기반 AI 코</description>
    </item>
    <item>
      <title>Terraform이란: 코드로 인프라를 구축·운영하는 IaC 도구</title>
      <link>https://th-yong.com/tech/terraform-introduction</link>
      <guid>https://th-yong.com/tech/terraform-introduction</guid>
      <pubDate>Tue, 24 Feb 2026 00:00:00 GMT</pubDate>
      <description>1. 개요 Terraform은 프로그램 코드로 인프라(Infrastructure)를 구축·변경·운영할 수 있게 해주는 오픈소스 IaC(Infrastructure as Code) 도구입니다. 콘솔에서 클릭(ClickOps)으로 서버/네트워크/보안 정책을</description>
    </item>
    <item>
      <title>automem: AI 에이전트에 지속적(Relational) 장기 메모리를 제공하는 그래프+벡터 하이브리드 메모리</title>
      <link>https://th-yong.com/tech/automem-durable-relational-memory</link>
      <guid>https://th-yong.com/tech/automem-durable-relational-memory</guid>
      <pubDate>Fri, 06 Feb 2026 00:00:00 GMT</pubDate>
      <description>1. 개요 automem은 AI 에이전트에게 세션 간에도 유지되는 지속적(durable)이며 관계형(relational) 장기 메모리를 제공하는 오픈소스 메모리 시스템입니다. 단순히 “비슷한 문장”을 찾는 수준을 넘어, 대화/결정에서 추출된 개체(En</description>
    </item>
    <item>
      <title>HMLR: Agentic AI를 위한 장기 기억 시스템</title>
      <link>https://th-yong.com/tech/HMLR</link>
      <guid>https://th-yong.com/tech/HMLR</guid>
      <pubDate>Mon, 26 Jan 2026 00:00:00 GMT</pubDate>
      <description>HMLR(Hierarchical Memory Lookup &amp; Routing)은 AI 에이전트를 위한 상태 인식 장기 기억 시스템이다. 기존의 단순 벡터 RAG 방식이 처리하지 못하는 시간적 충돌, 다중 홉 추론, 교차 토픽 기억 문제를 해결하기 위해 </description>
    </item>
    <item>
      <title>Agent-Lightning 분석</title>
      <link>https://th-yong.com/tech/agent-lightning-analysis</link>
      <guid>https://th-yong.com/tech/agent-lightning-analysis</guid>
      <pubDate>Tue, 23 Dec 2025 00:00:00 GMT</pubDate>
      <description>기존 에이전트 코드를 수정하지 않고 강화학습과 프롬프트 최적화를 적용하는 것이 특징</description>
    </item>
  </channel>
</rss>