AI커뮤니티

자랑/공유

인기 태그 #rag#RAG#opensource#agents#llm#에이전트#임베딩#비용#GPU#오픈소스#커뮤니티#serving
  1. 12
    추천
    이미지 링크
    사내 매뉴얼 2천 페이지를 RAG로 묶어서 챗봇을 만들었습니다. 임베딩은 bge-m3, 서빙은 vLLM으로 했고, 정확도는 사내 QA 기준 87% 나옵니다.
    자랑/공유 @hanbit · 7시간 전 · 댓글 1 #RAG#사내#배포
  2. 9
    추천
    이미지 링크
    인터넷이 안 되는 환경에서 LLM+RAG를 돌려야 하는 프로젝트였는데, Ollama + 로컬 임베딩으로 구성했습니다. 오프라인 배포에 관심 있으신 분들께 도움이 되길.
    자랑/공유 @sean · 9시간 전 · 댓글 1 #오프라인#RAG#사례
  3. 8
    추천
    이미지 링크
    매일 아침 AI 관련 뉴스 5개를 요약해서 텔레그램으로 보내주는 봇을 만들었습니다. DeepSeek로 요약하고 키워드로 분류하니 꽤 유용하네요.
    자랑/공유 @nova · 11시간 전 · 댓글 0 #텔레그램#뉴스#봇
  4. 7
    추천
    이미지 링크
    매주 LLM 논문 하나씩 리뷰하는 블로그를 6개월째 운영 중입니다. 방문자가 꾸준히 늘고 있고, 이메일 구독도 300명 넘었네요. 꾸준함이 답이긴 하네요.
    자랑/공유 @polyglot · 13시간 전 · 댓글 1 #블로그#논문#기록
  5. 11
    추천
    이미지 링크
    지식베이스 검색 시간이 하루 2시간 → 10분으로 줄었습니다. 초기엔 환각이 문제였는데, 인용 소스 표시를 강제하니 신뢰도가 확 올라갔어요.
    자랑/공유 @alexus · 18시간 전 · 댓글 1 #KB#RAG#생산성
  6. 9
    추천
    보안 때문에 전부 온프레미스로. bge-m3 + Qwen2.5-14B로 구축했고, 사용자 만족도 4.2/5 나왔어요. 후기 공유합니다.
    자랑/공유 @mlpark · 23시간 전 · 댓글 1 #사내#챗봇#온프레미스
  7. 8
    추천
    금융권 고객사에 설치했습니다. 인터넷 완전 차단 환경에서도 임베딩+검색+생성 전부 동작. 자세한 내용은 글에.
    자랑/공유 @sean · 1일 전 · 댓글 1 #폐쇄망#RAG#배포
  8. 7
    추천
    모든 데이터는 기기에 저장됩니다. 주제, 기분, 실행 항목이 담긴 주간 다이제스트. 2주 만에 사용자 400명!
    자랑/공유 @kai · 1일 전 · 댓글 1 #local-first#journal#product
  9. 9
    추천
    Ticket deflection이 한 달 만에 0%에서 38%로 증가했습니다. 아키텍처와 우리가 저지른 실수들을 공유합니다.
    자랑/공유 @alexus · 1일 전 · 댓글 1 #support#rag#case-study
  10. 11
    추천
    L4 한 장으로 Qwen2.5-7B 서빙. 처리량 만족스럽습니다. 양자화 없이도 충분하네요.
    자랑/공유 @mlpark · 1일 전 · 댓글 2 #vLLM#서빙#GPU
  11. 6
    추천
    Scrapes 20 sources, dedupes, summarizes, posts at 8am. 1.2k subscribers now. AMA about the pipeline.
    자랑/공유 @rai · 1일 전 · 댓글 1 #telegram#bot#newsletter
  12. 6
    추천
    주 1회 LLM 논문 리뷰. 누적 12만 뷰 달성. 글쓰기 + 실험 재현을 병행하니 배우는 게 두 배네요.
    자랑/공유 @hanbit · 1일 전 · 댓글 0 #블로그#논문#리뷰
  13. 7
    추천
    동일 7B 모델로 비교했는데 p50은 비슷하고 p99는 TRT-LLM이 더 안정적이었습니다. 수치 표로 정리했어요.
    자랑/공유 @gptkim · 1일 전 · 댓글 1 #vLLM#TensorRT#벤치마크
  14. 8
    추천
    Replaced a custom fine-tune with a generic 8B + good retrieval. Cheaper to maintain, better scores. Data point for you.
    자랑/공유 @vectorlee · 1일 전 · 댓글 1 #rag#finetune#case-study
  15. 4
    추천
    링크
    Worth a read — some genuinely interesting verticals (last-mile routing, predictive maintenance).
    자랑/공유 @rai · 1일 전 · 댓글 1 #startups#ml#vertical
  16. 6
    추천
    일반적인 이름으로 회의에 참여해서, 메모를 남기고 우리 위키에 게시합니다. 메모가 나타나기 전까지 사람들은 그 존재를 잊어버리죠. 전사에는 아주 작은 로컬 모델을 사용합니다. 한 번은 "이거 합법이냐"는 질문이 나온 적이 있는데, HR이 괜찮다고 해서 출시합니다.
    자랑/공유 @Marcus Obediah · 1일 전 · 댓글 0 #meetings#notes#side-project
  17. 5
    추천
    Visual node editor for agent pipelines with live tracing. MIT licensed, would love feedback.
    자랑/공유 @polyglot · 1일 전 · 댓글 0 #opensource#agents#playground
  18. 5
    추천
    매주 토요일 2시간. 30명에서 시작해 지금 12명의 코어 멤버로. 꾸준함의 가치를 배웠습니다.
    자랑/공유 @nova · 1일 전 · 댓글 1 #스터디#모임#커뮤니티
  19. 6
    추천
    Scrapes arXiv + HN, summarizes with an LLM, posts as threads. 200 stars in a week. AMA.
    자랑/공유 @rai · 1일 전 · 댓글 2 #newsletter#automation
  20. 8
    추천
    Our org has 10k pages of policy PDFs nobody reads. Built a RAG bot with citations that links back to the exact page. Support queries dropped noticeably and compliance is happy because everything is traceable. Side project that became the team's favorite tool.
    자랑/공유 @Jennifer Chen · 2일 전 · 댓글 0 #rag#chatbot#showcase
  21. 7
    추천
    Collected 3 years of my blog posts and newsletter, fine-tuned a 7B with LoRA. The style imitation is uncanny — my wife thought I wrote the drafts. Uses: first drafts of posts I'm too lazy to start. It's my ghostwriter now.
    자랑/공유 @Marcus Obediah · 4일 전 · 댓글 0 #finetuning#writing#lora
  22. 4
    추천
    Transcribe with Whisper, summarize with a fine-tuned small model, get a 5-bullet recap per episode. My backlog went from 40 episodes of guilt to a nice reading list. The summaries are honest enough that I know which episodes to actually listen to.
    자랑/공유 @Jennifer Chen · 3일 전 · 댓글 0 #podcast#whisper#summarization
  23. 6
    추천
    Workshop is loud, dusty, and has no wifi worth trusting. Pi 5 + a local model + a cheap mic array. It answers questions about my tools, sets timers, and reads specs out loud. Everything runs on-device and it cost under $200.
    자랑/공유 @Sofia Alvarez · 5일 전 · 댓글 0 #voice#offline#hardware
  24. 6
    추천
    링크
    Been dumping notes, highlights, and bookmarks into a local RAG setup for two years. Search is instant and surprisingly good. Open-sourced the whole thing so it's reproducible. Warning: your old notes are more embarrassing than you remember.
    자랑/공유 @David Kim · 6일 전 · 댓글 0 #second-brain#rag#opensource
  25. 7
    추천
    12B model on the family PC, a simple web UI. It explains math problems step by step instead of just giving answers. No accounts, no data leaving the house, and no monthly subscription. The kids actually use it.
    자랑/공유 @Raj Patel · 8일 전 · 댓글 0 #local-llm#education#family
  26. 5
    추천
    이 사이트와 비슷한데 반도체 뉴스에 특화된 거예요. 50개의 RSS 피드를 스크래핑하고, 임베딩으로 중복을 제거하고, LLM으로 요약한 다음, 제 읽기 습관에 따라 순위를 매겨요. 한 달째 아침마다 읽고 있는데 예전 피드 리더는 다시 열어본 적이 없어요.
    자랑/공유 @Emily Watson · 7일 전 · 댓글 0 #news#summarization#side-project
  27. 6
    추천
    내 git 커밋과 PR을 긁어와서 3개의 불릿 포인트로 요약해 스탠드업 전에 이메일로 보내준다. 아침에 30초면 끝난다. 우리 팀장은 내가 엄청 체계적인 사람인 줄 안다. 사실 난 그냥 단계만 늘린 게으름뱅이일 뿐인데.
    자랑/공유 @Anna Nowak · 9일 전 · 댓글 0 #automation#agents#humor
  28. 5
    추천
    It categorizes spending, flags weird charges, and answers questions like 'how much did I spend on coffee in March'. Accuracy is 90%+ on categories. Caveat: you have to trust it with your data, so it runs fully local. Worth it for the anxiety reduction alone.
    자랑/공유 @Tom Hall · 10일 전 · 댓글 0 #finance#local-llm#privacy