Tech News Google DeepMind researchers introduce new benchmark to improve LLM factuality, reduce hallucinations adminJanuary 10, 20250 Join our daily and weekly newsletters for the latest updates and exclusive content on industry-leading AI coverage. Learn More Hallucinations,…
Hackers News Identifying and Manipulating LLM Personality Traits via Activation Engineering adminDecember 31, 20240 Comments
Hackers News LLM Reasoning with Chain of Continuous Thought by Meta AI adminDecember 30, 20240 Facebook Twitter WhatsApp CopyCopied Introduction Large language models (LLMs) have demonstrated incredible reasoning abilities, penetrating an increasing number of domains…
Hackers News Making AMD GPUs competitive for LLM inference adminDecember 24, 20240 Aug 9, 2023 • MLC Community TL;DR MLC-LLM makes it possible to compile LLMs and deploy them on AMD GPUs…
Hackers News Offline Reinforcement Learning for LLM Multi-Step Reasoning adminDecember 23, 20240 Comments
Tech News IBM wants to be the enterprise LLM king with its new open-source Granite 3.1 models adminDecember 19, 20240 Join our daily and weekly newsletters for the latest updates and exclusive content on industry-leading AI coverage. Learn More IBM…
Hackers News Apple collaborates with NVIDIA to research faster LLM performance adminDecember 18, 20240 In a blog post today, Apple engineers have shared new details on a collaboration with NVIDIA to implement faster text…
Hackers News Extend (YC W23) is hiring engineers to build LLM document processing adminDecember 18, 20240 Comments
Hackers News Ask HN: Examples of Agentic LLM Systems in Production? adminDecember 16, 20240 Now that everybody and their mother are fuzzing in social media about LLM agents and agentic LLM systems (or something),…
Tech News New LLM optimization technique slashes memory costs up to 75% adminDecember 13, 20240 Universal Transformer Memory uses neural networks to determine which tokens in the LLM’s context window are useful or redundant.Read More
Hackers News Task-Specific LLM Evals that Do & Don’t Work adminDecember 9, 20240 If you’ve ran off-the-shelf evals for your tasks, you may have found that most don’t work. They barely correlate with…