Open Source LLM

AOMP 19.0-0 Released For AMD's OpenMP Offloading Compiler

AMD on Thursday published AOMP 19.0-0 as the newest version of their LLVM/Clang downstream compiler focused on delivering the latest OpenMP device offloading support for their Radeon GPUs and Instinct accelerators.

The version bump with AOMP 19.0-0 comes as a result of now tracking the latest LLVM/Clang 19.0 upstream Git given the recent stable release of LLVM/Clang 18. With this AOMP 19.0-0 release they are now building against the ROCm 6.0
Continue reading...

Open Source LLM

Open-source AI models released by Tokyo lab Sakana founded by former Google researchers - Reuters

Open-source AI models released by Tokyo lab Sakana founded by former Google researchers Reuters

Open Source LLM

OpenAI's GPT-5, their next-gen foundation model is coming soon

Serving tech enthusiasts for over 25 years.
TechSpot means tech analysis and advice you can trust .

A hot potato: ChatGPT, the chatbot that turned machine learning algorithms into a new gold rush for Wall Street speculators and Big Tech companies, is merely a "storefront" for large language models within the Generative Pre-trained Transformer (GPT) series. Developer OpenAI is now readying yet another upgrade for the technology.

OpenAI is busily working on GPT-5, the next generation of the company's multimodal large language

Continue reading...

Open Source LLM

(PR) Supermicro Launches Three NVIDIA-Based, Full-Stack, Ready-to-Deploy Generative AI SuperClusters

Supermicro, Inc., a Total IT Solution Provider for AI, Cloud, Storage, and 5G/Edge, is announcing its latest portfolio to accelerate the deployment of generative AI. The Supermicro SuperCluster solutions provide foundational building blocks for the present and the future of large language model (LLM) infrastructure. The three powerful Supermicro SuperCluster solutions are now available for generative AI workloads. The 4U liquid-cooled systems or 8U air-cooled systems are purpose-built and designed for powerful LLM training performance, as

Open Source LLM

If you teach a chatbot how to read ASCII art, it will teach you how to make a bomb

Serving tech enthusiasts for over 25 years.
TechSpot means tech analysis and advice you can trust .

In context: Most, if not all, large language models censor responses when users ask for things considered dangerous, unethical, or illegal. Good luck getting Bing to tell you how to cook your company's books or crystal meth. Developers block the chatbot from fulfilling these queries, but that hasn't stopped people from figuring out workarounds.

University researchers have developed a way to "jailbreak" large language models like

Continue reading...

Open Source LLM|Apple Silicon

Apple reveals AI model that can interpret photos and count objects

Serving tech enthusiasts for over 25 years.
TechSpot means tech analysis and advice you can trust .

What just happened? Apple has been slow to adopt generative AI, but that might be changing with the introduction of MM1, a multimodal large language model capable of interpreting both image and text data. This functionality could potentially be included in the company's next generation of handsets and services although there are also rumors of Apple integrating Google's Gemini AI.

Apple researchers have developed MM1, a new approach for

Continue reading...

Open Source LLM|Apple Silicon

Report: Apple to Use Google's Gemini AI for iPhones

In the world where the largest companies are riding the AI train, the biggest of them all—Apple—seemed to stay quiet for a while. Even with many companies announcing their systems/models, Apple has stayed relatively silent about the use of LLMs in their products. However, according to Bloomberg, Apple is not pushing out an AI model of its own; rather, it will license Google's leading Gemini models for its iPhone smartphones. Gemini is Google's leading AI model with three variants: Gemini

Open Source LLM

(PR) MemVerge and Micron Boost NVIDIA GPU Utilization with CXL Memory

MemVerge, a leader in AI-first Big Memory Software, has joined forces with Micron to unveil a groundbreaking solution that leverages intelligent tiering of CXL memory, boosting the performance of large language models (LLMs) by offloading from GPU HBM to CXL memory. This innovative collaboration is being showcased in Micron booth #1030 at GTC, where attendees can witness firsthand the transformative impact of tiered memory on AI workloads.

Charles Fan, CEO and Co-founder of MemVerge, emphasized the critical importance of overcoming the bottleneck

Open Source LLM

(PR) Phison Announces Strategic Partnerships Deploying aiDAPTIV+ at NVIDIA GTC 2024

Phison Electronics, a leading provider of NAND controllers and storage solutions, today announced aiDAPTIV+ partnerships with ASUS, Gigabyte, MAINGEAR, and MediaTek. At GTC 2024, Phison and partners will demonstrate aiDAPTIV+, a hybrid hardware and software large language model (LLMs) fine-tune training solution that enables small and medium-sized businesses (SMBs) to process and retain local control of their sensitive machine learning (ML) data.

Foundational training of LLMs gives a broad understanding of language but aiDAPTIV+ enables

Open Source LLM

(PR) MAINGEAR Introduces PRO AI Workstations Featuring aiDAPTIV+ For Cost-Effective Large Language Model Training

MAINGEAR, a leading provider of high-performance custom PC systems, and Phison, a global leader in NAND controllers and storage solutions, today unveiled groundbreaking MAINGEAR PRO AI workstations with Phison's aiDAPTIV+ technology. Specifically engineered to democratize Large Language Model (LLM) development and training for small and medium-sized businesses (SMBs), these ultra-powerful workstations incorporate aiDAPTIV+ technology to deliver supercomputer LLM training capabilities at a fraction of the cost of traditional AI training servers.

As the demand for large-scale generative