본문으로 건너뛰기

Training (학습/파인튜닝)

Training (학습/파인튜닝) 관련 AI 뉴스를 한국어로 요약해 드립니다.

이 카테고리에서 얻을 수 있는 정보

Training (학습/파인튜닝) 카테고리는 글로벌 AI/ML 뉴스에서 해당 주제와 직접 연결된 핵심 기사만 모아 빠르게 탐색할 수 있는 허브입니다. 태그별 빈도와 최신성을 함께 확인해 흐름을 파악하고, 관심 주제의 상세 기사로 바로 이동할 수 있습니다.

관심 태그 추가
LoRA228GRPO160RL110RLHF99PPO90RLVR33DPO32XGBoost30SFT27AdamW18DINO17QLoRA14PEFT12SAC12Behavior Cloning11DQN11Active Learning9Continual Learning9Random Forest9Deep Learning8Physical AI8Backpropagation7Gradient Descent7On-policy distillation7Bayesian Inference6Multimodal RL6OPSD6Variational Inference6Contrastive learning5cross entropy5Distributed Training5Federated Learning5Fine-tuning5Graph Neural Network5Imitation Learning5OPD5Rectified Flow5Self-play5Sim-to-Real5Test-Time Training5DAPO4DMD4Predictive Coding4Self-Supervised Learning4DAgger3Ensemble Methods3Group Relative Policy Optimization (GRPO)3HACPO3HG-DAgger3Pre-training3REINFORCE3RLAIF3SDPO3Self Forcing3STDP3TD33Transfer Learning3UMAP3ALS2Barlow Twins2BetaZero2Causal RL2Compositional Learning2Curriculum Learning2Data Augmentation2DDP2DIAYN2Domain Adaptation2double descent2Dropout2EMA2ERD2Franca2GAIL2GiGPO2GloVe2Go-Explore2Gradient Boosting2GRAM2GSPO2gVisor2Hamiltonian-SMT2Hebbian Learning2IMG Sign2InfoNCE2Instruction Tuning2Inverse Dynamics2MAP-Elites2Maximum Likelihood Estimation2MEMIT2Neuro-Symbolic2Online Learning2Optogenetics2ORB2Physics-Informed ML2POMCP2PRM2Prompt Tuning2Pseudo-Labeling2QAT2REPO2reward hacking2Reward Hypothesis2Reward Model2RFT2Ridge regression2RLMF2Self-Distillation2Self-Refine23DreamBooth18-fold data augmentation1A2C1A3C1ABC1Ability-aware Environment Selection (AES)1ACC1Adaptive Gradient Clipping1ADMIRE1Adversarial Evolution1AdvFD1AES-SIV1AGE1Agentic Data Flow1Agentic RL1Agentic Training1AgentOPSD1AGI safety course1ALTER1AMVL1Anchor-Align1Answer-Backtracked Credit Assignment (ABC)1ARC-Forcing1Arithmetic Repetition Complexity1Astrolabe1AtomiMed1Attention Z-Reg1autoresearch@home1AWR1Back-translation1BEACON1BERTMap1Bilinear Game1Behavioral Multi-Token Prediction (BMTP)1Boosted Control Function1Boosting1BRAID1C3-GP1CAA1CAPI1CAST1Catastrophic Forgetting1Causal Analysis1Causal-rCM1CCPL1CE1Centroid1CLaaS1CLeaD1CLIP-style objective1Coconut-style1Complete(d)P1Concept Data Attribution1Consistency Distillation1ConstrainedZero1Continual RL1Continuous Learning1Continuous Online Learning1Contrastive Decoding Diffing1ControlG1convolve1Coreset Selection1CoRT1Counterfactual Perturbation Consistency1CoVe1CPDN1CPR1Cross-Layer Value Routing (CLVR)1Cross-modal Distillation1Cycle-Consistency1DanceOPD1DAPD1DAPR1DART-SD1Data-Conditioned Method Planning1Data-Construction-Skill1Data Labeling1Dataset Enrichment1ddiff1Deep Delta Learning1Deep Clustering1DeepSearch-Evolve1DenoiseRL1Dense Panoramic Ray-Conditioning (DPRC)1DICE-RL1Diffusion-based Semantic Compression (DiSCo)1DiffusionNFT1DiffusionOPSD1diffusion teacher1Digital Red Queen1Direct-OPD1Disentanglement1Distribution Matching1DO-ALL1DOPD1dOPSD1DPE1DP-FedGD1DPPR1DP-SGD1DreamBooth1Dream Rehearsal1Dr. GRPO1Drop-Then-Recovery (DTR)1DRQ1DSDR1Dueling DDQN1dynamic_select1Elastic-net1EML1Empirical Likelihood1Empirical Risk Minimization1End-to-End Learning1Epsilon-greedy1Equivariant Optimal-Transport Flow Matching1ERM1ESPO1Evidential Learning1Evolutionary Strategies1EWC1Exact-match Caching1Expectation Maximization1Faithful-RFT1FastOPD1FastSVERL1f-divergence1FD-Loss1FedAvg1FedNova1FedProx1FedUMM1FIM1FlashSAC1FlowMirror1Flux-GS1Focal Loss1FOCUS1From Scratch1FullFT1Fused Lasso1GBNF1GDPO1Generative Causal Testing (GCT)1Genetic Programming1GenFirst1GenomeDiff1Geometry-Aware Memory Augmentation (GMA)1GeoPT1GFPO1GIFT1GP-C-LUCB1Gradient Accumulation1Gradient-Based Connections1Group Based RL1Grouped Cross-Entropy (GCE)1He Initialization1HeRA1Hierarchical Difficulty Curriculum (HDC)1HIL1Holistic Data Scheduler1Homoscedastic Uncertainty Weighting1IC-LoRA1ICRL1ID Representation Forcing1Implicit Differentiation1Implicit Error Counting (IEC)1InfoPO1InnerZoom1INSIGHT1Instruction Selection1Interdisciplinary Doctoral Program in Statistics (IDPS)1J-Zero1Keyframe Evidence Memory1KNN1Laplacian Visual Prompting1Latent Lookahead1LESS1LIMEN1LIPPAX1LK Loss1LLM-teacher Distillation1LMO1Local Extra SGD1Local Learning Rule1LOEO1Logit Mixing1LoKR1LongForcing1LoopRPT1LOPE1LoRS1Louvain1LSVI-UCB1MALA1Manana1MANCE1Manifold Bandits1MAPD1Masked Boundary Modeling1Masked Depth Modeling1Matryoshka Learning1MAXIS Loss1MaxSim1MeanFlowNFT1Medprompt1MegaTrain1MemLearner1Memory-R11Meta-Learning1MET-D1Midtraining1MIM1MindForge1MIPI1MIPU1Mixture-of-Kittens1MixUp1ML Modelling1MMProLong1Momentum1Monte Carlo Dropout1MOPD1MRPO1MSE1MSFT1MultAttnAttrib1Multi4D1Multiprobe Grid1Multi-Task Learning1MultiTF1MV-Forcing1Native Factorized Weights (NFW)1NCE1NEAT1Nested Learning1Next-Token Prediction1NOCs1Noise Contrastive Estimation1Nonlinear Regression1NoRA1Numeric Gradient1OAR-Flow1OGAR1OmniBoost1OmniDPO1Open-Reasoner-Zero1OPID1OPSA1Optimization1OraRL1OrthoReg1OSCAR1OSFB1Oversampling1PAC Learning1Parallel-RL1Parameter-Efficient Fine-Tuning1Pass-Conditioned Reading1PatternRL1PDM1PEPO1Perceptual Hashing1Perceptual Flow Matching1PGD1Stability Metric Φ1PICA1PipeDream-2BW1PIPO1PIRL1PivotRL1POET-X1Policy Conditioning1Policy Gradient1Policy Iteration1Policy Optimization1PopuLoRA1Post-training1PowerCool1PRA1PRA-GRPO1Prior Selection1Progress Advantage1Progressive Training1Protein Language Model1Prototypical Contrastive Learning1Proxy Prompt Reinforcement Learning1PTR1Pulsed Learning1PuzzleMix1Q-Synth Module1Quantization-Aware Healing1Quantization-Aware Training1Quantum Machine Learning1RAD-GRPO1RAHA1RayPE1RCORE1rDPO1ReGFT1Relay-OPD1Remember‑R11Requential Coding1Resource2Skill1Reverse KL1Reverse-KL Divergence1reward shaping1RFE1Riemannian Flow Matching1RLHEV1RL-Index1RLinf1RLOO1RLSD1RLSVR1RLVF1R-MFT1RoboTTT1RoCEv21rollback reward1ROPD1RP-SFT1SA-DMD1SAGE-GRPO1SaMer1Sample Weight Decay1SAMPO1SA-MRPO1SAO1SAR1SEED1Selective Learning1Self-Distilled Reasoning (SDR)1Self-Evolution1Self-Flow1Self Gradient Forcing1SELFCI1Self-OPD1Self-Patching1SetFit1SGBM1SHA-2561Shapley values1ShortOPD1Signed MaxSim1SimpleOPD1SimPO1Simulated Environment1SKILL01Skill Self-Play1SMOTE1Soft Actor-Critic1SoftDTW1Soufflé Datalog1SPA1SPAM1Sparse Upcycling1SPIFFE1SRA1SRL1SRM-LoRA1SSKD1SSync1Stable LatentMoE1Stochastic Gradient Descent1StudentSim1Subliminal Learning1Super Sensing1SVERL1SVRL1SWE-RL1SWIM1synthetic dynamics1TACO1TacSL1Task-Aware Knowledge Compression (TAKC)1TASKER1TBPTT1TD3-BC1TDM-R11Teacher Forcing1Test-Time Learning1Tinylora1TN-GSPO1Token-Superposition Training1TOP-D1TORC1Training Dynamics1Training-Free1Transductive Learning1Transliteration1TREK1T-Rex1TRIAGE1TTPO1TTRL1TurboDiffusion1TurnOPD1UniMatch V21u-OPSD1URLVR1USAF1VAD1Value Iteration1Verbalizable Representations1VESPO1VibeWorlding-Gym1Visual Pretraining1VR-GRPO1V-Zero1WarpSAC1WavFlow1Weight Space Learning1Working Memory Depth Recurrence1world rehearsal1XKD-Dial1zElo1

자주 묻는 질문

Training (학습/파인튜닝) 카테고리에는 어떤 기준으로 기사가 포함되나요?

기사의 핵심 주제와 엔티티 태그를 기준으로 자동 분류되며, 관련성이 낮거나 중복된 항목은 우선순위가 낮게 조정됩니다.

태그 옆 숫자는 무엇을 의미하나요?

각 태그가 연결된 누적 기사 수를 의미합니다. 숫자가 클수록 해당 주제의 노출 빈도가 높은 편입니다.

최신 기사만 빠르게 보려면 어떻게 해야 하나요?

하단의 모든 뉴스 보기 링크를 통해 필터된 피드로 이동하면 최신순으로 관련 기사를 연속 탐색할 수 있습니다.