<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Thedduro</title><link>https://thedduro.github.io/</link><description>Recent content on Thedduro</description><generator>Hugo</generator><language>ko-kr</language><lastBuildDate>Mon, 21 Sep 2026 14:22:11 +0900</lastBuildDate><atom:link href="https://thedduro.github.io/index.xml" rel="self" type="application/rss+xml"/><item><title>Spark cache(), 어디에 붙여야 효과가 있을까?</title><link>https://thedduro.github.io/posts/spark-cache-when-to-use/</link><pubDate>Mon, 21 Sep 2026 14:22:11 +0900</pubDate><guid>https://thedduro.github.io/posts/spark-cache-when-to-use/</guid><description>&lt;p&gt;API 요청 로그에서 잘못된 데이터를 걸러냈다. 이제 유효한 요청이 몇 건인지 세고, 경로별 평균 응답 시간도 구하려고 한다. 전처리한 DataFrame 하나를 두 작업에서 사용하면 되니, 전처리도 한 번만 실행될 것처럼 보인다.&lt;/p&gt;</description></item><item><title>소개</title><link>https://thedduro.github.io/about/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://thedduro.github.io/about/</guid><description>&lt;p&gt;데이터 사이언스 전공으로 시작해 B2B 컨설팅 기업에서 데이터 분석 업무를 수행했습니다. 분석 결과가 장기적인 비즈니스 자산으로 이어지기까지의 어려움을 경험하면서, 데이터를 분석하는 일에서 데이터를 계속 활용할 수 있는 시스템을 만드는 일로 관심을 넓혔습니다.&lt;/p&gt;</description></item></channel></rss>