Posts

Showing posts with the label Self-Improvement

DeepSeek

Introduction:  Nvidia神話不再?!AI行業大變天⚠今次真係東升西降☀ 實測Deepseek R1: 淺白解說其特別之處 | 為什麼會吸引全球高度關注? DeepSeek震撼美股! The Engineering Unlocks Behind DeepSeek | YC Decoded Development: Deep-Agent / R1-V DeepSeek R1 GAVE ITSELF a 200% Speed Boost - Self-Evolving LLM Did DeepSeek R1 Invent a Language Humans CAN'T Understand? DeepSeek R1 Cloned for $30?! PhD Student STUNNING Discovery Politics: 城寨新聞 III 27/1/2025: (9:50) 城寨新聞 I 29/1/2025: 中共人工智能  (21:52) 城寨咖啡室: 2/2/2025: DeepSeek 〈蕭若元:蕭氏新聞台〉2025-01-28 〈蕭若元:蕭氏新聞台〉2025-01-30 《蕭若元:蕭氏新聞台》2024-01-31 《蕭若元》2025-02-02 《蕭氏新聞台》2025-02-04 What DeepSeek is Teaching us about U.S. National Security: A breakdown DeepSeek遭多地政府禁用 改變國運的突破?DeepSeek引爆中國社群媒體 Alexandr Wang Is DeepSeek lying? Scale AI CEO Alexandr Wang on U.S.-China AI race Legal and Copyright Issues OpenAI says DeepSeek used its models illegally Microsoft and OpenAI investigate   Deepseek懶人包 AI War DeepSeek異軍突起 能否助中國勝出美中科技戰?- BBC News 中文

AI Self-Improvement

Research Papers  Research efforts have explored methods to enable LLMs to refine their outputs through self-reflection and iterative processes. A study titled "Large Language Models Can Self-Improve" demonstrated that LLMs could enhance their reasoning abilities by generating high-confidence, rationale-augmented answers for unlabeled questions and fine-tuning themselves using these self-generated solutions. arxiv.org Another approach, detailed in the paper "SELF: Self-Evolution with Language Feedback," involves a framework where LLMs engage in self-reflection and refinement, iteratively improving their performance without human intervention. This method draws inspiration from human learning processes, emphasizing the potential for LLMs to evolve through self-assessment and feedback. arxiv.org

Artificial Super Intelligence (ASI)

Theory  Artificial Super Intelligence (ASI) is imminent - Cognitive Hyper Abundance is coming (Jan 19, 2025) AI Models Close to ASI DEEPSEEK DROPS AI BOMBSHELL: A.I Improves ITSELF Towards Superintelligence (BEATS o1) (Jan 20, 2025)