Author: Steve

  • The Unplugged AI: Could Local LLMs Avert the Coming Datacenter Crisis?

    A silent revolution is brewing on our desktops and laptops, one that could have a profound impact on the future of artificial intelligence and the very infrastructure of the internet. As the world marvels at the capabilities of large language models (LLMs) accessible through online chatbots, a growing trend of running these models locally is emerging, powered by tools like Ollama. This shift from remote, cloud-based AI to personal, localized processing may be more than just a niche for tech enthusiasts; it could be a crucial step in mitigating a predicted global shortage of datacenter capacity.

    The insatiable demand for AI is pushing our global datacenter infrastructure to its limits. Generative AI is driving a massive surge in electricity consumption, with some forecasts predicting a 165% increase in power demand from data centers by 2030 compared to 2023.[1] This explosive growth is creating an “insatiable demand for power that will exceed the ability of utility providers to expand their capacity fast enough,” according to Gartner, which predicts that 40% of existing AI data centers will be operationally constrained by power availability by 2027.[2] This trajectory has led to headlines about AI data centers consuming as much electricity as small cities and has companies like Meta, Google, and Microsoft spending billions on datacenter buildouts.[3][4]

    At the heart of this issue is the energy-intensive nature of both training and running the massive LLMs that power popular online services.[5][6] Training these models is a monumental task, requiring immense computational power over extended periods. But the “silent killer,” as some have noted, is inference – the energy consumed every time a user asks a question or generates text.[6] An AI-powered search, for instance, can consume 10 to 30 times more energy than a traditional one.[6] This constant, large-scale demand is a primary driver of the predicted datacenter crunch.

    However, a different paradigm is gaining traction. The rise of powerful, open-source LLMs, coupled with user-friendly tools like Ollama, is making it increasingly practical for individuals and businesses to run these models on their own hardware.[7][8][9] This “local AI revolution” offers a compelling alternative to a future dominated by a few centralized AI providers.[9]

    The benefits of running LLMs locally are numerous. Users gain greater privacy and control over their data, as sensitive information doesn’t need to be sent to third-party servers.[7][10][11][12] This is a critical consideration for industries like healthcare and finance.[8][12] Furthermore, local LLMs eliminate ongoing API costs, offering a potentially more cost-effective solution for high-usage applications after the initial hardware investment.[7] The ability to customize and fine-tune models for specific needs is another significant advantage, freeing users from the constraints of vendor ecosystems.[7]

    But what about the energy consumption of these local models? While it may seem counterintuitive, running an LLM for inference on a personal computer is surprisingly efficient. The key difference lies in the nature of the workload. Unlike the sustained, high-power demands of training or large-scale commercial inference, local usage is typically “bursty.”[3] A user sends a request, the hardware processes it for a few seconds, and then returns to an idle state.[3] One user, running a local LLM server in a high-cost energy market, noted that the actual power impact on their home setup was so small it was “barely worth thinking about.”[3] The primary cost associated with local LLMs is the initial hardware purchase, not the ongoing electricity bill.[3]

    This is not to say that local LLMs will completely replace their cloud-based counterparts. The future of AI is likely to be a “pluralistic” one, with both centralized and decentralized models coexisting to serve different needs.[10] Large, general-purpose models in the cloud will continue to be essential for complex, large-scale tasks.[10] However, for a significant portion of everyday AI interactions, specialized and efficient local models can provide a powerful and private alternative.

    The widespread adoption of local LLMs could have a profound global impact. By offloading a substantial portion of AI inference from centralized data centers to individual devices, we could significantly reduce the strain on our global power grids. This decentralized approach could help to democratize access to AI, empowering individuals and small businesses to leverage its capabilities without relying on large tech monopolies.[13]

    The path to a more decentralized AI future is not without its challenges. The initial hardware investment can be a barrier for some, and deploying and maintaining local models still requires a degree of technical expertise.[14] However, the rapid advancements in hardware and the continuous development of user-friendly software are steadily lowering these barriers.[15]

    The conversation around the future of AI has been rightly focused on its transformative potential. But as we stand on the cusp of a potential datacenter crisis, it’s time to also consider the sustainability of our approach. The move towards local LLM usage, facilitated by tools like Ollama, presents a compelling vision for a more distributed, resilient, and ultimately, a more sustainable AI ecosystem. It’s a future where the power of artificial intelligence resides not just in the cloud, but on the devices we use every day, potentially averting a global infrastructure bottleneck in the process.

    Sourceshelp

    1. goldmansachs.com
    2. gartner.com
    3. xda-developers.com
    4. d-matrix.ai
    5. aimultiple.com
    6. medium.com
    7. ipsofactointeractif.ca
    8. senseisrl.it
    9. fatihbattal.com.tr
    10. artiba.org
    11. busyday.co.uk
    12. neilsahota.com
    13. bitforgedynamics.com
    14. intradatech.com
    15. aclu.org
  • Addressing the Looming Crisis in Clinical Diagnostics: A Call for Strategic Investment

    Once in a while, it’s good to step back from the daily grind of lab operations, EMR integrations, and security reviews to reflect on the fundamental infrastructure that underpins modern healthcare: the clinical laboratory workforce.

    The diagnostic revolution—from Next Generation Sequencing to advanced molecular testing—has delivered incredible tools, but these tools are only as effective as the highly-skilled professionals running the assays and managing the resulting data. Right now, that critical foundation is cracking.

    The Diagnostic Bottleneck and Its Impact

    The statistics are sobering, though perhaps unsurprising to anyone managing a clinical lab. Vacancy rates for clinical laboratory professionals are exceptionally high—some reports suggest they are close to 25%. This isn’t merely a human resources challenge; it’s a systemic threat to patient care because the diagnostic work performed in the lab informs nearly 70% of a physician’s medical decisions.

    When a lab is understaffed, the consequences are immediate and severe:

    • Operational Strain: Existing staff face burnout and high-stress environments.
    • Patient Impact: Turnaround times lengthen, directly delaying diagnoses and treatment plans.
    • Future Challenges: As the population ages and the demand for complex, high-volume testing (like genomics) grows, this shortage will become a critical constraint on the entire healthcare system.

    A Strategic Solution: The Path to Pipeline Resilience

    It’s encouraging to see a national effort to address this crisis with a long-term, strategic solution. The recently re-introduced bipartisan legislation, the Medical Laboratory Personnel Shortage Relief Act, offers a concrete, multi-pronged approach to reinforcing the workforce pipeline.

    This bill correctly focuses on two major, intertwined levers that are essential for successful recruitment and retention:

    1. Direct Financial Incentives (Loan Forgiveness): By including lab personnel in the National Health Service Corps (NHSC), professionals would become eligible for student loan forgiveness programs. This is a powerful, instant incentive that makes a lab career path more attractive and helps offset significant educational debt, which is a major barrier to entry and retention.
    2. Capacity Expansion (Training Grants): The current bottleneck often lies in the limited capacity of accredited university programs and the availability of qualified teaching faculty. The proposed federal grant program aims to directly fund institutions and hospitals to expand their training programs, support necessary internships, and increase the number of qualified individuals prepared to enter the field.

    Key Takeaway for Lab Professionals

    As technical leaders and program managers, we understand that every complex system requires a stable, well-resourced foundation to achieve scalable success. Investing in the lab workforce is not an optional expense; it is a vital, strategic necessity for maintaining the quality and capacity of our healthcare system.

    For lab leaders, this legislation highlights the areas where advocacy matters most: supporting programs that enable loan repayment and those that directly fund capacity expansion in educational programs. Without a robust and growing supply of highly trained lab professionals, our advancements in precision medicine and complex diagnostics simply cannot be sustained.

  • Unraveling the Potential: Incretin Mimetics in the Context of Public Health

    Unraveling the Potential: Incretin Mimetics in the Context of Public Health

    As we traverse the broader terrain of biomedical science, the recently emerging class of drugs, called incretin mimetics, stands out with a potential to significantly influence public health, particularly in the United States. These drugs, including Glucagon-like peptide-1 (GLP-1) inhibitors such as liraglutide (Saxenda), semaglutide (Ozempic and Wegovy), and dulaglutide (Trulicity), ushers in a new direction in the management of both diabetes and obesity.

    The action of incretin mimetics can be visualized as being akin to stepping on nature’s accelerator – enhancing the beneficial effects of incretin hormones. They bind to the GLP-1 receptors and induce glucose-dependent insulin release, earning their stripes as effective antihyperglycemics. As an added benefit, they suppress appetite, slow the emptying of our stomachs, and reduce glucagon secretion, effectively reducing the sharp elevation of blood glucose levels typically seen post meals.

    Standing at the intersection of the American diabetes and obesity epidemic – with over 34 million people living with diabetes and an estimated 42.4% adult obesity prevalence – the role of incretin mimetics could be transformative.

    The vigor of incretin mimetics such as Saxenda, Ozempic, Trulicity, and Wegovy is not merely noteworthy, it might indeed prove critical. In clinical trials, these potential behemoths have triumphed, regulating blood sugar and inducing weight loss. An unexpected but welcome adjunct has been their significant reduction of major cardiovascular events, a worrisome complication typically associated with type 2 diabetes.

    When brought into the context of public health in the U.S., the potential benefits offered by incretin mimetics is somewhat short of astonishing. They directly address two major health crises—diabetes and obesity—and by virtue of their cardiovascular benefits, these agents promise to alleviate the financial strains currently borne by the healthcare system. Reduced hospital admissions, lessened healthcare costs, and an improvement in patient quality of life paint a picture of a healthier future.

    As the future unfolds, we can hope that research and development continue keeping pace, refining these therapeutic agents to offer potentially improved clinical outcomes. However, this optimism comes with a caveat: these results need to be consistent over the long term. If they are, we may be on the brink of a decisive turning point in clinical medicine and public health.

    Unlocking access to these treatments to a broader patient population presents a contemporary challenge. Should this be achieved, we might make significant strides toward democratizing healthcare, stimulate progress, and reinforce our initiative to improve public health. The age of incretin mimetics is dawning. It symbolizes not just hope, but a promise for a healthier United States.

  • Strategic Cloud Adoption in Healthcare: Leveraging Existing Infrastructure for Innovation

    Strategic Cloud Adoption in Healthcare: Leveraging Existing Infrastructure for Innovation

    In the dynamic landscape of healthcare IT, the conversation around cloud adoption isn’t just about embracing new technologies—it’s about leveraging existing infrastructure to drive innovation. Many healthcare organizations have already made significant investments in big data centers or private cloud environments, and understanding how to augment these investments with public cloud offerings like AWS and Google Cloud is paramount.

    Here’s how this nuanced approach changes the calculation of cloud adoption:

    1. Hybrid Cloud Strategy: Instead of viewing cloud adoption as an all-or-nothing proposition, healthcare organizations can adopt a hybrid cloud strategy that combines the strengths of both private and public cloud environments. This approach allows organizations to leverage existing investments in on-premises infrastructure while harnessing the scalability and flexibility of the public cloud for specific use cases.
    2. Optimizing Workloads: Not all workloads are created equal, and healthcare organizations must carefully evaluate which workloads are best suited for the cloud. While mission-critical applications with stringent security and compliance requirements may remain on-premises, less sensitive workloads such as development and testing environments or data analytics projects can be migrated to the cloud for cost savings and agility.
    3. Data Governance and Compliance: Healthcare organizations operate within a highly regulated environment, and compliance with regulations such as HIPAA is non-negotiable. When considering cloud adoption, organizations must ensure that their chosen cloud provider adheres to industry-specific compliance standards and provides robust data governance capabilities to protect patient information.
    4. Cost Optimization: While the cloud offers scalability and flexibility, it’s essential to carefully manage costs to avoid overspending. Healthcare organizations should conduct thorough cost-benefit analyses to determine the most cost-effective deployment model for each workload, taking into account factors such as data transfer costs, storage requirements, and usage patterns.
    5. Integration and Interoperability: Seamless integration with existing systems and interoperability with third-party applications are critical factors in successful cloud adoption. Healthcare organizations must ensure that their chosen cloud provider offers robust integration capabilities and supports industry-standard protocols to facilitate interoperability across disparate systems.

    Practical Takeaways:

    1. Assess Your Current Infrastructure: Take stock of your organization’s existing infrastructure, including big data centers and private cloud environments, to identify opportunities for optimization and integration with public cloud offerings.
    2. Develop a Comprehensive Cloud Strategy: Develop a holistic cloud strategy that aligns with your organization’s goals, taking into account factors such as workload requirements, compliance considerations, and cost implications.
    3. Engage Stakeholders: Cloud adoption is a multifaceted endeavor that requires buy-in from stakeholders across the organization, including IT teams, compliance officers, and executive leadership. Engage stakeholders early and often to ensure alignment and mitigate potential roadblocks.
    4. Continuously Evaluate and Iterate: The healthcare landscape is constantly evolving, and so too should your cloud strategy. Continuously evaluate the effectiveness of your cloud deployment model and iterate as needed to stay ahead of the curve.

    By taking a nuanced approach to cloud adoption that leverages existing infrastructure while embracing the benefits of public cloud offerings, healthcare organizations can unlock new opportunities for innovation, collaboration, and improved patient care.

  • The Unspoken Pillar of Cloud Computing: Standardization

    The Unspoken Pillar of Cloud Computing: Standardization

    In the fast-paced world of technology, few innovations have reshaped our digital landscape as profoundly as cloud computing. It has revolutionized the way individuals and organizations store, process, and access data, ushering in an era of unparalleled scalability, flexibility, and efficiency. Yet, amidst the discussions of its essential characteristics, one aspect often overlooked is standardization—a fundamental trend shaping the very fabric of cloud computing.

    Cloud computing has been characterized by several indispensable traits, including on-demand self-service, broad network access, resource pooling, rapid elasticity, and measured service. These characteristics have formed the cornerstone of cloud architecture, enabling businesses to scale their operations dynamically, optimize resource utilization, and drive innovation. However, amidst these widely acknowledged attributes, standardization stands out as a silent yet indispensable force driving the evolution of cloud technology.

    Standardization in cloud computing entails the establishment of uniform protocols, interfaces, and practices across different cloud environments, whether they are public, private, or hybrid. It facilitates interoperability, seamless integration, and consistent management of resources, regardless of their underlying infrastructure. This means that users can deploy and manage applications with ease, irrespective of the specific cloud platform they choose, leading to greater flexibility and efficiency in IT operations.

    The significance of standardization becomes even more pronounced in today’s increasingly complex computing landscape, characterized by diverse deployment models, heterogeneous environments, and multi-cloud strategies. As organizations embrace a mix of public cloud services, private clouds, and on-premises infrastructure, the need for standardized approaches becomes paramount to streamline operations, mitigate complexity, and ensure optimal performance across the board.

    Moreover, standardization fosters innovation by promoting an open ecosystem where vendors, developers, and users can collaborate more effectively. By adhering to common standards and frameworks, stakeholders can leverage interoperable tools, share best practices, and accelerate the development of new technologies and solutions. This collaborative ethos not only drives efficiency but also empowers organizations to harness the full potential of cloud computing to address complex business challenges and drive digital transformation.

    Looking ahead, the trajectory of cloud computing suggests that standardization will continue to play a pivotal role in shaping the future of IT infrastructure. As the industry matures and evolves, we can anticipate a convergence towards a unified management layer, where disparate cloud environments seamlessly integrate into a cohesive ecosystem. This holistic approach to cloud management promises to simplify IT operations, enhance agility, and lower costs, ultimately delivering a more frictionless experience for users.

    In conclusion, while cloud computing has been lauded for its transformative impact on the digital landscape, standardization emerges as a silent yet indispensable trend driving its evolution. By fostering interoperability, collaboration, and innovation, standardization lays the groundwork for a more cohesive and efficient cloud ecosystem. As organizations navigate the complexities of modern IT environments, embracing standardized practices will be essential to unlock the full potential of cloud computing and pave the way for a seamless, interconnected future.

  • Pharmacogenomics: What you need to know

    Pharmacogenomics: What you need to know

    Pharmacogenomics is an emerging field that is revolutionizing the way healthcare is delivered. By leveraging advancements in genomics and molecular biology, pharmacogenomics enables healthcare professionals to tailor treatments and medications to an individual’s genetic makeup. This means that medications are more targeted and effective and can be tailored to a patient’s specific needs.

    Pharmacogenomics can help to reduce the risk of adverse drug reactions, as it identifies a patient’s genetic makeup and allows healthcare professionals to select the most appropriate medications for that individual. This can help to reduce the time and money spent on ineffective treatments, as well as the risk of serious side effects.

    Pharmacogenomics also has the potential to reduce the time and expense associated with clinical trials. By using genetic data to determine the most likely clinical outcomes for a particular drug, researchers can design more effective clinical trials and identify patients who may respond better to a particular drug.

    One of the most exciting applications of pharmacogenomics is in the area of personalized medicine. By looking at a patient’s individual genetic profile, doctors can identify the treatments and medications that are most likely to be effective, as well as those that may cause adverse reactions. This personalized approach to medicine has the potential to improve health outcomes and reduce overall healthcare costs.

    As this technology continues to evolve and become more widely used, pharmacogenomics will play a significant role in the way healthcare is delivered. It has the potential to reduce the amount of time and money spent on ineffective treatments, improve health outcomes, and reduce healthcare costs. With its many advantages, pharmacogenomics is certainly a technology worth watching.

  • Sandworm: A new era of cyberwar and the hunt for the Kremlin’s most dangerous hackers

    Sandworm: A new era of cyberwar and the hunt for the Kremlin’s most dangerous hackers

    I was pleasantly surprised by Andy Greenberg’s dramatic recount of recent cybersecurity events, which included valuable insight into the progression of cyberwarfare tactics to include real-world impact on infrastructure. 

    It was helpful to have a high-level review of these events in a style that was highly engaging. I found myself still amazed and horrified to hear the gory details of the NotPetya attack and the damage unleashed within the US as well as in Russia, as well as the power system attacks leading up to in in Georgia and Ukraine. It’s clear that the era of cyberwarfare has already arrived.

    Greenberg’s storytelling ability makes this a captivating work, even for those already familiar with the subjects he covers, but it also stands on its own. He takes time to not only wade into the dubious field of attribution, but to digest and reevaluate the relationship of subgroups, eventually seeming to settle on a collaborative model.

    This was an entertaining and informative book, which I would feel comfortable recommending to people interested in this area, no matter what their background and experience.

  • This is How they Tell me The World Ends: The Cyberweapons Arms Race

    Once in a while, it’s good to take a step back from the daily grind of security and governance to stop and think about where this is all going. Nicole Perlroth’s title seemed like an enticing opportunity for just such an interlude and, after shopping around I was able to get early access to this work and was satisfied with what it had to offer. She uses her narrative to guide the reader through the shady underworld of cybercrime, focusing on the growing zero day exploit market and providing some cool insights into its relatively banal and perhaps well-intentioned origins, leading into an overview of the current state, which is not quite world-ending but perhaps world-changing.

    While most cybercriminals still take advantage of the old standard vectors, (i.e., misconfigured or unpatched systems and uneducated users), zero day exploits seem to offer special powers, enabling malicious actors to invade fully patched systems without alerting users. This has made them an irresistable tool to state actors with deep pockets, who are able to use them as the perfect surveillance tool for intelligence gathering. State actors have proven to be willing to pay large bounties for zero day exploits, which they secretly retain until they are needed for intelligence operations. Even large companies (i.e., the big 4) are simply unable to keep up with governments who are willing to pay 6 figures for critical exploits. This creates an incentive for independent bug hunters to keep their findings private, ultimately leading to less secure computing environments for all users. 

    Up until now, over-eager attribution to ‘State Actors’ and APTs has been something of a clichè, it seems these type of actors may play a more dominant role in security incidents, as stockpiling zero-days has proven to be a valuable strategy by such actors, even including US and Israeli intelligence agencies. To parse the idea of ‘world-ending’ in a bit more detail, we can expect to encounter an increasing number of undisclosed vulnerabilities as time goes on. In Rumsfeldlish, these are ‘unknown unknowns’–issues that we didn’t even know we should worry about. In reality, zero-days are negative-days, since the impact could have already been encountered for an unknown amount of time prior to their discovery. If your system has been specifically targeted by a state actor using a zero-day, then there really isn’t much you could directly do to defend against it. However, it does raise the importance of general DR planning, including failover systems, site recovery strategies and offline backups. National borders also have heightened importance in this type of scenario, since a state actor would essentially have unlimited power to act within their own borders, so if your business strategy leverages cloud infrastructure, it might be worth considering how your global footprint could reduce the potential impact of rogue state actors in other parts of the world.

    As I’ve hopefully illustrated, I am glad Nicole has created this insightful work, which is both an engaging narrative and a discussion starter security people.