Mistral’s Impact on the AI Landscape

Mistral models, recognized for their open nature, have quickly become leaders among open-source LLMs. The company burst onto the scene by releasing capable open-weights models, Mistral 7B and Mixtral 8×7B. Mistral is partnering with Microsoft Azure to provide their models through Azure AI Studio and Azure Machine Learning in addition to their own platform andContinue reading “Mistral’s Impact on the AI Landscape”

Navigating the Future of AI in the Creative Industries

Subscribe • Previous Issues The Impact of Text-to-Video Models on Video Production Sora is a large-scale AI system from OpenAI capable of generating high-fidelity videos up to a minute long using just text prompts. It employs neural networks (“diffusion transformer architecture”) to acquire a diverse range of video simulation capabilities that could profoundly impact the entertainment andContinue reading “Navigating the Future of AI in the Creative Industries”

Five Reasons Developers Should Be Excited About Gemini

My recent experiments with the Gemini API  have yielded encouraging outcomes. Up to this point, my access has been limited to Gemini 1.0 Pro. Although I harbored no illusions about it outperforming GPT-4, my experiences have largely been favorable. While the 1.0 Pro version isn’t perfect and sometimes produces perplexing results, I am confident theseContinue reading “Five Reasons Developers Should Be Excited About Gemini”

Early Thoughts on Claude 3

Anthropic’s next-generation foundation model, Claude 3, boasts three variants: Haiku, Sonnet, and Opus. Each tackles cognitive tasks with unparalleled expertise, catering to a wide range of needs.  You can find the Model Card here. At the heart of Claude 3’s design is a core emphasis on enhanced intelligence across various domains. Whether it’s knowledge acquisition,Continue reading “Early Thoughts on Claude 3”

Managing the Risks and Rewards of Large Language Models

Large language models (LLMs) have exploded in capability and adoption over the past couple years. They can generate human-like text, summarize documents, translate between languages, and even create original images and 3D designs based on text descriptions. Companies remain highly bullish on LLMs, with most either actively experimenting with or already partially implementing the technologyContinue reading “Managing the Risks and Rewards of Large Language Models”

From Supervised Fine-Tuning to Online Feedback

Over the last 9 months, my usage of general-purpose language models like OpenAI’s API has decreased as I’ve learned to leverage open-source models fine-tuned for specific tasks. Anyscale’s user-friendly Fine Tuning service has accelerated this transition by making it easy to craft accurate, efficient custom models.  Despite the initial investment in creating labeled datasets, theContinue reading “From Supervised Fine-Tuning to Online Feedback”

How Generative AI is Transforming Healthcare

Subscribe • Previous Issues Generative AI in Healthcare: Beyond the Horizon of Modern Medicine Whenever a new technology emerges, I like to explore its application across various sectors, especially those that are highly regulated such as financial services and healthcare. These sectors, with their exacting standards and stringent regulations, provide a robust framework for evaluating the maturityContinue reading “How Generative AI is Transforming Healthcare”

Generative AI’s Impact on Healthcare

The healthcare sector is enormously complex, requiring advanced tools to unlock innovation. As highlighted in our latest report, generative AI brings transformative potential across healthcare, from accelerating drug discovery to optimizing hospital operations. Explore the Report: Clinical Support and Documentation: Dive into how AI is revolutionizing the way clinicians interact with patient data, enhancing decision-makingContinue reading “Generative AI’s Impact on Healthcare”

localllm and the Promise and Pitfalls of Running LLMs Locally

localllm is an open-source framework that aims to democratize the use of large language models (LLMs) by enabling their efficient operation on local CPUs. This circumvents the need for expensive and scarce GPUs. It provides developers with an easy way to access state-of-the-art quantized LLMs from Hugging Face through a simple command-line interface. localllm canContinue reading “localllm and the Promise and Pitfalls of Running LLMs Locally”

AMD’s Expanding Role in Shaping the Future of LLMs

In my recent exploration of emerging hardware options for Large Language Models (LLMs), AMD’s offerings have emerged as particularly promising. In this analysis, I delve deeper into the factors that position AMD GPUs favorably for leveraging the growth of LLMs and Generative AI. These factors range from performance and efficiency gains in demanding AI tasksContinue reading “AMD’s Expanding Role in Shaping the Future of LLMs”