Tren
dar
대시보드
논문
뉴스
GitHub
AI 소식
KO
EN
로그인
AI 소식
논문
대시보드
뉴스
GitHub
“language model”
이 키워드와 관련된 논문 · GitHub · 뉴스를 한곳에 모았습니다.
논문
12
전체 →
Semantic Scholar
자연어·LLM
인용 927
의사 수준의 의료 질문 답변을 위한 대규모 언어 모델
Toward expert-level medical question answering with large language models
Semantic Scholar
자연어·LLM
인용 530
추론·에이전트 성능과 효율성을 동시에 높인 오픈 LLM
DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
Semantic Scholar
자연어·LLM
인용 407
모든 가중치를 1.58비트로 표현하는 초저비용 LLM
The Era of 1-bit LLMs: All Large Language Models are in 1.58 Bits
Semantic Scholar
자연어·LLM
인용 269
진단 대화를 위한 AI 시스템 AMIE 개발
Towards conversational diagnostic artificial intelligence
OpenAlex
자연어·LLM
인용 1.4K
대규모 언어 모델의 발전과 활용에 대한 종합적 조사
A Survey of Large Language Models
Semantic Scholar
자연어·LLM
인용 1.7K
지속적 수집으로 오염 없는 코드 LLM 평가 벤치마크
LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code
Semantic Scholar
자연어·LLM
인용 1.6K
ChatGLM-4, GPT-4에 필적하는 중국어·영어 LLM 시리즈
ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
Semantic Scholar
자연어·LLM
인용 1.3K
의미 수준의 불확실성 측정으로 LLM 환각 탐지
Detecting hallucinations in large language models using semantic entropy
Semantic Scholar
자연어·LLM
인용 890
검색 증강 생성(RAG)과 LLM의 결합 동향을 종합적으로 분석한 서베이
A Survey on RAG Meeting LLMs: Towards Retrieval-Augmented Large Language Models
Semantic Scholar
자연어·LLM
인용 568
LLM이 언어 대신 연속 잠재 공간에서 추론하도록 학습
Training Large Language Models to Reason in a Continuous Latent Space
Semantic Scholar
멀티모달
인용 568
강화학습으로 멀티모달 대규모 언어모델의 추론 능력 향상
Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Semantic Scholar
자연어·LLM
인용 554
LLM의 수학적 추론 능력 한계를 밝힌 GSM-Symbolic 벤치마크
GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models
GitHub
5
전체 →
Python
★ 31.5K
sgl-project/sglang
Python
★ 11K
mrexodia/ida-pro-mcp
컴퓨터비전
Python
★ 3K
데이터 중심 전략으로 구현된 최신 비전-언어 모델 패밀리
NVlabs/Eagle
Jupyter Notebook
★ 21.7K
guidance-ai/guidance
Jupyter Notebook
★ 27.6K
HandsOnLLM/Hands-On-Large-Language-Models
뉴스
11
전체 →
Hacker News
자연어·LLM
▲ 42
소형 언어 모델의 임베딩 응집 현상을 완화하는 분산 손실 함수
Dispersion loss counteracts embedding condensation in small language models
Hacker News
자연어·LLM
▲ 7
Anthropic, 언어 모델의 글로벌 워크스페이스 개념 제시
Anthropic: A global workspace in language models
Hacker News
산업·기업
▲ 5
일반 목적 LLM이 전문화된 임상 AI 도구보다 의료 벤치마크에서 우수
General-purpose large language models outperform specialized clinical AI tools
Hacker News
안전·보안
▲ 4
LLM이 선호하는 가상 인물 쌍이 웹과 학술 출판에 미치는 영향
AI language models have favorite names, and we mapped them
Anthropic
자연어·LLM
▲ 0
Anthropic, 언어 모델의 글로벌 워크스페이스 구조 연구 발표
A global workspace in language models - Anthropic
mit_tr
산업·기업
▲ 0
LLM 병목 해결 주장하는 스타트업, 서브쿼드러틱
A startup claims it broke through a bottleneck that’s holding back LLMs
theverge
제품·출시
▲ 0
OpenAI, 브로드컴과 협업해 자체 AI 추론 칩 '할라피뇨' 공개
OpenAI reveals its first AI processor: Jalapeño
mit_tr
▲ 0
The Download: Claude’s inner workings and OpenAI’s “super app”
mit_tr
▲ 0
Anthropic found a hidden space where Claude puzzles over concepts
techcrunch
▲ 0
Your gaming data could be the secret to AGI, according to this Bezos-backed startup
techcrunch
▲ 0
Why this CEO thinks video games make better training data than the internet