A machine learning approach where a model learns by trial and error, receiving rewards or penalties for its actions.
मॉडेल चाचणी व चुकीतून शिकते, त्याच्या कृतींसाठी बक्षीस किंवा दंड मिळवत, अशी मशीन लर्निंग पद्धत.
Reinforcement learning trains a system through feedback on outcomes rather than labelled examples, a technique used in some stages of modern language model training.
रिइन्फोर्समेंट लर्निंग प्रणालीला लेबल केलेल्या उदाहरणांऐवजी निकालांवरील अभिप्रायातून प्रशिक्षित करते, आधुनिक भाषा मॉडेल प्रशिक्षणाच्या काही टप्प्यांत वापरले जाणारे तंत्र.
Contact: fin100x.ai@gmail.com · Fin100X.AI Pvt. Ltd., Maharashtra, India
AI outputs are advisory; final decision authority rests with authorised government officials.