Practical Generative AI Applications and LLMs

Practical Generative AI Applications and LLMs

Recent advances in generative AI, exemplified by LLMs like Stable Diffusion and ChatGPT, have created significant industry hype. Generative AI involves creating new media (such as text or images) by analyzing massive datasets to deduce and mimic existing patterns, a process driven by probabilistic and stochastic modeling. While models like GPT can produce humanlike text, they operate as language prediction models rather than utilizing true reasoning (AGI), which means they often "stumble over facts," produce inconsistent results, and struggle with basic tasks like multiplication, leading to "hallucinations". To leverage these tools effectively, prompt engineering is necessary—this "subtle art" involves providing clear, specific instructions, setting a system context or persona, and potentially using examples to coax a useful result from the AI. When integrating AI via the stateless Completions API, developers must manually maintain conversation state by sending the entire history with each request, often summarizing older messages to manage token costs. More robust applications can utilize GPT Functions (Tools) to allow the model to intelligently call external functions—avoiding expensive model retraining—to access live or proprietary data. Alternatively, to query custom data using natural language, facts can be converted into high-dimensional vectors called embeddings and compared using cosine similarity against user queries, often managed in a database like Postgress with PG Vector. Finally, the newer Assistants API simplifies the development of domain-specific helpers by automatically managing message history and context compaction, and uniquely, when referencing uploaded knowledge files (like a lease document), it provides specific references or footnotes detailing where the answer was found.


Ref: https://www.youtube.com/watch?v=OxHw_u45h7M&list=PL03Lrmd9CiGey6VY_mGu_N8uI10FrTtXZ&index=18

Denne episoden er hentet fra en åpen RSS-feed og er ikke publisert av Podme. Den kan derfor inneholde annonser.

Episoder(144)

The 7 Skills You Need to Build AI Agents

The 7 Skills You Need to Build AI Agents

As AI agents become more capable, the skills needed for AI jobs are shifting. Bri Kopecki breaks down the 7 skills you need to move from prompt engineering to full agent engineering, including system ...

11 Aug 19min

What is LangChain?

What is LangChain?

LangChain became immensely popular when it was launched in 2022, but how can it impact your development and application of AI models, Large Language Models (LLM) in particular. In this video Martin Ke...

5 Aug 20min

LangChain vs LangGraph

LangChain vs LangGraph

Get ready for a showdown between LangChain and LangGraph, two powerful frameworks for building applications with large language models (LLMs.) Master Inventor Martin Keen compares the two, taking a lo...

30 Jul 16min

RAG vs Agentic AI

RAG vs Agentic AI

Agentic AI and RAG are redefining how LLMs think and act 🤖. Live from TechXchange in Orlando, Martin Keen & Cedric Clyburn unpack how vector databases, data integration, and context engineering enabl...

23 Jul 21min

RAG's Evolution

RAG's Evolution

How did search evolve into agentic AI? Sam Anthony explains RAG's evolution, from simple retrieval to adaptive systems powered by LLMs. Learn how semantic search, hybrid retrieval, and AI agents enabl...

16 Jul 13min

AI Agent Skills

AI Agent Skills

We're all using AI agents, but they still lack the procedural knowledge real work needs. Martin Keen explains how agent skills, LLMs, RAG, and MCP help agents follow workflows, automate tasks, and mak...

9 Jul 24min

MCP vs. RAG

MCP vs. RAG

How do AI agents learn and take action? Live from TechXchange in Orlando, Melissa Hadley breaks down how MCP and RAG help large language models connect to data — one to retrieve knowledge, the other t...

2 Jul 20min

RAG vs Fine-Tuning vs Prompt Engineering

RAG vs Fine-Tuning vs Prompt Engineering

How do AI chatbots deliver better responses? Martin Keen explains RAG 🛠️, fine-tuning , and prompt engineering methods that extend knowledge, refine responses, and build domain expertise. Learn how t...

25 Jun 20min

Populært innen Fakta

fastlegen
dine-penger-pengeradet
relasjonspodden-med-dora-thorhallsdottir-kjersti-idem
foreldreradet
treningspodden
jakt-og-fiskepodden
rss-kunsten-a-leve
rss-strid-de-norske-borgerkrigene
mikkels-paskenotter
hverdagspsyken
sinnsyn
fryktlos
gravid-uke-for-uke
rss-var-forste-kaffe
rss-sarbar-med-lotte-erik
laringsmiljo-i-skole-og-barnehage-uis-podkast
rss-orjasater
rss-impressions-2
kvallm
lrerrommet