{"id":846,"date":"2023-09-25T08:11:17","date_gmt":"2023-09-25T08:11:17","guid":{"rendered":"https:\/\/tbekk.com\/devstream\/?p=846"},"modified":"2023-10-05T09:45:18","modified_gmt":"2023-10-05T09:45:18","slug":"weekly-ai-and-nlp-news-september-18th-2023","status":"publish","type":"post","link":"https:\/\/tbekk.com\/devstream\/2023\/09\/25\/weekly-ai-and-nlp-news-september-18th-2023\/","title":{"rendered":"Weekly AI and NLP News \u2014 September 18th 2023"},"content":{"rendered":"\n<h2 class=\"wp-block-heading\">Stable Audio, Mixture-of-Experts LLMs, and LLMs for compiler optimization<\/h2>\n\n\n\n<hr class=\"wp-block-separator has-text-color has-light-gray-color has-alpha-channel-opacity has-light-gray-background-color has-background is-style-wide\"\/>\n\n\n\n<ul class=\"wp-block-list\">\n<li><em><strong>Link:<\/strong><\/em> <a href=\"https:\/\/medium.com\/nlplanet\/weekly-ai-and-nlp-news-september-18th-2023-3e128fbda17d\"><em>NLPlanet<\/em><\/a><\/li>\n\n\n\n<li><strong><em>Author:<\/em><\/strong> <a href=\"https:\/\/medium.com\/@chiusanofabio94?source=post_page-----3e128fbda17d--------------------------------\"><em>Fabio Chiusano<\/em><\/a><\/li>\n\n\n\n<li><strong><em>Publication Date:<\/em><\/strong> <em>Sept. 18, 2023<\/em><\/li>\n<\/ul>\n\n\n\n<hr class=\"wp-block-separator has-text-color has-light-gray-color has-alpha-channel-opacity has-light-gray-background-color has-background is-style-wide\"\/>\n\n\n\n<p id=\"2277\">Here are your weekly articles, guides, and news about NLP and AI chosen for you by&nbsp;<a href=\"https:\/\/www.nlplanet.org\/\" rel=\"noreferrer noopener\" target=\"_blank\">NLPlanet<\/a>!<\/p>\n\n\n\n<h1 class=\"wp-block-heading\" id=\"c969\">\ud83d\ude0e News From The Web<\/h1>\n\n\n\n<ul class=\"wp-block-list\">\n<li><a href=\"https:\/\/www.stableaudio.com\/\" rel=\"noreferrer noopener\" target=\"_blank\">Stable Audio<\/a>. London-based startup Stability AI, known for its AI model Stable Diffusion, has launched Stable Audio, an AI model that can generate high-quality commercial music with more control over synthesized audio.<\/li>\n\n\n\n<li><a href=\"https:\/\/www.reuters.com\/technology\/google-nears-release-ai-software-gemini-information-2023-09-15\/\" rel=\"noreferrer noopener\" target=\"_blank\">Google nears release of AI software Gemini, The Information reports<\/a>. Google is close to launching Gemini, an advanced language model that will rival GPT4. It is currently in the early testing phase and offers a range of functionalities including chatbots, text summarization, and code writing assistance.<\/li>\n\n\n\n<li><a href=\"https:\/\/www.theregister.com\/2023\/09\/12\/openai_copyright_lawsuits\/\" rel=\"noreferrer noopener\" target=\"_blank\">Pulitzer Prize winner and others sue OpenAI<\/a>. Pulitzer-winning novelist Michael Chabon and other writers are suing OpenAI for copyright infringement, claiming that the datasets used to train ChatGPT contain copyrighted content. OpenAI argues that its language learning models are protected by \u201cfair use,\u201d igniting discussions on AI and copyright law in the field.<\/li>\n\n\n\n<li><a href=\"https:\/\/techcrunch.com\/2023\/09\/13\/adobes-firefly-generative-ai-models-are-now-generally-available-get-pricing-plans\/\" rel=\"noreferrer noopener\" target=\"_blank\">Adobe\u2019s Firefly generative AI models are now generally available<\/a>. Adobe has released commercially available generative AI models in their Creative Cloud, including a standalone web app called Firefly. The new \u201cgenerative credits\u201d system controls user interactions with Firefly\u2019s AI models, with each click on \u2018generate\u2019 using one credit.<\/li>\n\n\n\n<li><a href=\"https:\/\/www.theverge.com\/2023\/9\/8\/23863943\/roblox-ai-chatbot-assistant-ai-rdc-2023\" rel=\"noreferrer noopener\" target=\"_blank\">Roblox\u2019s new AI chatbot will help you build virtual worlds<\/a>. Roblox\u2019s 2023 Developers Conference introduced the Roblox Assistant, a new conversational AI tool designed to assist creators in developing more immersive virtual experiences. This tool enables creators to easily generate virtual environments and implement basic gameplay behaviors. However, it will not be accessible until the end of this year or early next year.<\/li>\n<\/ul>\n\n\n\n<h1 class=\"wp-block-heading\" id=\"16a5\">\ud83d\udcda Guides From The Web<\/h1>\n\n\n\n<ul class=\"wp-block-list\">\n<li><a href=\"https:\/\/magazine.sebastianraschka.com\/p\/llm-training-rlhf-and-its-alternatives\" rel=\"noreferrer noopener\" target=\"_blank\">LLM Training: RLHF and Its Alternatives<\/a>. This content provides a guide on alternatives to Reinforcement Learning from Human Feedback (RLHF). It presents five different approaches with corresponding research papers. These alternatives include Constitutional AI, The Wisdom of Hindsight, Direct Preference Optimization, Reinforced Self- Training, and Scaling Reinforcement Learning from Human Feedback with AI Feedback.<\/li>\n\n\n\n<li><a href=\"https:\/\/www.salesforce.com\/news\/press-releases\/2023\/09\/07\/ai-usage-research\/\" rel=\"noreferrer noopener\" target=\"_blank\">New AI Usage Data Shows Who\u2019s Using AI \u2014 and Uncovers a Population of \u2018Super-Users\u2019<\/a>. Generative AI is experiencing steady growth in usage, with nearly half of the population utilizing it and a third using it daily. Younger generations, particularly Gen Z and Millennials, are the \u201csuper users\u201d of generative AI, with 65% of them embracing the technology and trusting its decision- making guidance.<\/li>\n\n\n\n<li><a href=\"https:\/\/huggingface.co\/blog\/overview-quantization-transformers\" rel=\"noreferrer noopener\" target=\"_blank\">Overview of natively supported quantization schemes in \ud83e\udd17 Transformers<\/a>. Quantization schemes in Transformers like BitsandBytes and Auto-GPTQ offer ways to run large models on smaller devices. BitsandBytes is user-friendly and supports various models, while Auto-GPTQ excels in text generation speed but may result in lower quality. Both schemes can minimize performance degradation in larger models according to the Open-LLM leaderboard.<\/li>\n\n\n\n<li><a href=\"https:\/\/txt.cohere.com\/validating-llm-outputs\/\" rel=\"noreferrer noopener\" target=\"_blank\">Validating Large Language Model Outputs<\/a>. LLMs are powerful but can produce inconsistent results. Validating outputs is essential for reliable and accurate applications. Guardrails AI is a useful open-source package that improves LLM outputs by providing structural and quality assurances.<\/li>\n\n\n\n<li><a href=\"https:\/\/pub.towardsai.net\/create-a-self-moderated-commentary-system-with-langchain-and-openai-406a51ce0c8d\" rel=\"noreferrer noopener\" target=\"_blank\">Create a Self-Moderated Commentary System with LangChain and OpenAI<\/a>. This guide explains the process of creating a self-moderated commentary system using OpenAI and LangChain. It involves two models: one generates a response to user input, and the other analyzes and modifies the response before publishing it.<\/li>\n<\/ul>\n\n\n\n<h1 class=\"wp-block-heading\" id=\"74d8\">\ud83d\udd2c Interesting Papers and Repositories<\/h1>\n\n\n\n<ul class=\"wp-block-list\">\n<li><a href=\"https:\/\/github.com\/IBM\/ModuleFormer\" rel=\"noreferrer noopener\" target=\"_blank\">IBM releases MoE LLMs<\/a>. IBM has just released MoE LLMs, including models with 4B and 8B parameters. These models offer comparable computational efficiency to dense models with fewer parameters. They have been trained on a large dataset and utilize the ModuleFormer architecture.<\/li>\n\n\n\n<li><a href=\"https:\/\/arxiv.org\/abs\/2309.07062\" rel=\"noreferrer noopener\" target=\"_blank\">Large Language Models for Compiler Optimization<\/a>. Researchers have developed a powerful transformer model that optimizes LLVM assembly code for code size. The model outperforms baselines and demonstrates excellent code reasoning abilities, achieving a 3% reduction in instruction counts compared to compiler output. It generates compilable code 91% of the time and perfectly emulates the compiler\u2019s output 70% of the time.<\/li>\n\n\n\n<li><a href=\"https:\/\/arxiv.org\/abs\/2309.04564\" rel=\"noreferrer noopener\" target=\"_blank\">When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale<\/a>. Researchers have found that perplexity is a more effective method than complex scoring techniques for pruning pretraining data for language models.<\/li>\n\n\n\n<li><a href=\"https:\/\/arxiv.org\/abs\/2309.05519\" rel=\"noreferrer noopener\" target=\"_blank\">NExT-GPT: Any-to-Any Multimodal LLM<\/a>. NExT-GPT is an any-to-any multimodal language model that can process and generate content in various modalities such as text, images, videos, and audio. It achieves this by utilizing already-trained encoders and decoders, with minimal parameter tuning required.<\/li>\n\n\n\n<li><a href=\"https:\/\/github.com\/microsoft\/promptflow\" rel=\"noreferrer noopener\" target=\"_blank\">Microsoft releases Prompt Flow<\/a>. Microsoft has introduced Prompt Flow, a development suite for LLM-based apps. It offers a range of functionalities including creating executable workflows, debugging and iterating flows, evaluating flow quality and performance with larger datasets, integrating testing and evaluation into CI\/CD systems, and deploying flows to chosen serving platforms or app code bases easily.<\/li>\n\n\n\n<li><a href=\"https:\/\/arxiv.org\/abs\/2309.04269\" rel=\"noreferrer noopener\" target=\"_blank\">From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting<\/a>. A recent study introduces the \u201cChain of Density\u201d (CoD) prompting technique that generates dense summaries using GPT-4. By iteratively adding important entities without increasing the length, the resulting abstract summaries outperformed standard prompt summaries in terms of abstractive quality and reduced lead bias.<\/li>\n\n\n\n<li><a href=\"https:\/\/arxiv.org\/abs\/2309.07430\" rel=\"noreferrer noopener\" target=\"_blank\">Clinical Text Summarization: Adapting Large Language Models Can Outperform Human Experts<\/a>. LLMs have shown promising results in clinical text summarization tasks, surpassing human experts in terms of completeness and correctness. This research is the first to demonstrate LLMs outperforming humans in multiple clinical summarization tasks.<\/li>\n\n\n\n<li><a href=\"https:\/\/arxiv.org\/abs\/2309.03926\" rel=\"noreferrer noopener\" target=\"_blank\">Large-Scale Automatic Audiobook Creation<\/a>. New neural text-to-speech technology and automated parsing of e-books in the Project Gutenberg collection have resulted in the creation of over 5,000 open-license audiobooks, expanding the accessibility of this vast literature collection.<\/li>\n<\/ul>\n\n\n\n<p id=\"2881\">Thank you for reading! If you want to learn more about NLP, remember to follow&nbsp;<a href=\"https:\/\/www.nlplanet.org\/\" rel=\"noreferrer noopener\" target=\"_blank\">NLPlanet<\/a>. You can find us on&nbsp;<a href=\"https:\/\/www.linkedin.com\/company\/nlplanet\" rel=\"noreferrer noopener\" target=\"_blank\">LinkedIn<\/a>,&nbsp;<a href=\"https:\/\/twitter.com\/nlplanet_\" rel=\"noreferrer noopener\" target=\"_blank\">Twitter<\/a>,&nbsp;<a href=\"https:\/\/medium.com\/nlplanet\">Medium<\/a>, and our&nbsp;<a href=\"https:\/\/discord.gg\/zfC862H2dJ\" rel=\"noreferrer noopener\" target=\"_blank\">Discord server<\/a>!<\/p>\n","protected":false},"excerpt":{"rendered":"<p>Stable Audio, Mixture-of-Experts LLMs, and LLMs for compiler optimization Here are your weekly articles, guides, and news about NLP and AI chosen for you by&nbsp;NLPlanet! \ud83d\ude0e News From The Web&#8230; <a class=\"read-more-link\" href=\"https:\/\/tbekk.com\/devstream\/2023\/09\/25\/weekly-ai-and-nlp-news-september-18th-2023\/\">Read more &raquo;<\/a><\/p>\n","protected":false},"author":1,"featured_media":0,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[181,51,115,19,201,9],"tags":[27,20,302,40],"class_list":["post-846","post","type-post","status-publish","format-standard","hentry","category-ai-2","category-article","category-data-science","category-ml","category-nlp","category-news","tag-ai","tag-ml","tag-news","tag-nlp"],"_links":{"self":[{"href":"https:\/\/tbekk.com\/devstream\/wp-json\/wp\/v2\/posts\/846","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/tbekk.com\/devstream\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/tbekk.com\/devstream\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/tbekk.com\/devstream\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/tbekk.com\/devstream\/wp-json\/wp\/v2\/comments?post=846"}],"version-history":[{"count":1,"href":"https:\/\/tbekk.com\/devstream\/wp-json\/wp\/v2\/posts\/846\/revisions"}],"predecessor-version":[{"id":847,"href":"https:\/\/tbekk.com\/devstream\/wp-json\/wp\/v2\/posts\/846\/revisions\/847"}],"wp:attachment":[{"href":"https:\/\/tbekk.com\/devstream\/wp-json\/wp\/v2\/media?parent=846"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/tbekk.com\/devstream\/wp-json\/wp\/v2\/categories?post=846"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/tbekk.com\/devstream\/wp-json\/wp\/v2\/tags?post=846"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}