2401

rag-web-ui

An intelligent dialogue system that allows users to build custom Q&A services by combining their own document knowledge bases with LLMs using Retrieval-Augmented Generation.

2402

fastembed

A lightweight, fast Python library for generating text, image, and multimodal embeddings using ONNX Runtime to avoid heavy PyTorch dependencies.

2403

EvoAgentX

EvoAgentX is an open-source framework for building and automatically evolving LLM-based agents and workflows, featuring built-in evaluation and a rich library of tools.

2404

AnyCrawl

AnyCrawl is a high-performance web crawling and scraping toolkit that enables the extraction of structured JSON data from websites using LLMs, making web content LLM-ready.

2405

SimpleMem

A unified memory stack for LLM agents that uses semantically lossless compression to store and retrieve text and multimodal memories efficiently.

2406

GenerativeAIExamples

A collection of reference implementations and tutorials for building generative AI systems, RAG pipelines, and agentic workflows using the NVIDIA software ecosystem and NIM microservices.

2407

Town Square: Bringing Real-Time Human Presence to the Static Web

Town Square is an open-source widget that transforms websites into social spaces by visualizing current visitors as stick figures who can interact in real-time without accounts or permanent history.

2408

The Exploitarium Repository: Analysis of Mass-Dropped Undisclosed Vulnerabilities

An anonymous GitHub user has released 'exploitarium,' a consolidated archive of proof-of-concept exploits and vulnerability research targeting dozens of popular open-source projects, sparking debate over AI-driven bug hunting.

2409

Lealone

Lealone is a high-performance, self-evolving general agent that enables full-stack application and enterprise AI service development through natural language and SQL-like commands.

2410

holmesgpt

An open-source AI agent for SREs that automates production incident investigation and root cause analysis across any infrastructure stack.

2411

Asian AI Startups Launch Mythos-like Models Amid US Export Bans

Asian AI startups Sakana AI and 360 are launching frontier models like Fugu and Tulongfeng to fill the gap created by US export bans on Anthropic's Mythos and Fable 5.

2412

The Illusion of Digital Ownership: Why Physical Media Still Matters

Modern digital purchases are often revocable licenses rather than true ownership, making physical media and DRM-free alternatives the only reliable ways to preserve content.

2413

ai

A type-safe, provider-agnostic TypeScript SDK for building streaming chat, tool-calling agents, and multimodal AI applications across multiple JS frameworks.

2414

AdalFlow

AdalFlow is a PyTorch-like library for building and auto-optimizing LLM workflows, including chatbots, RAG, and agents, by replacing manual prompting with automated textual gradient descent.

2415

chonkie

A lightweight, high-performance text chunking library for RAG pipelines that provides diverse splitting strategies and seamless integrations with vector databases and embedding providers.

2416

OpenMemory

A cognitive memory engine for AI agents that provides long-term, multi-sector memory and temporal reasoning, moving beyond simple vector-based RAG.

2417

llm-graph-builder

A tool that uses LLMs and LangChain to transform unstructured data from various sources into structured Knowledge Graphs stored in Neo4j.

2418

LLM-Engineers-Handbook

A production-ready framework and codebase for building end-to-end LLM systems, covering training, RAG, and deployment on AWS.

2419

openagent

OpenAgent is an open-source personal AI assistant platform that combines LLMs, RAG, and autonomous agent loops into a single self-hostable binary for browser, shell, and office automation.

2420

ComfyUI-Copilot

An intelligent assistant for ComfyUI that automates workflow generation, debugging, and parameter tuning to streamline AI image generation development.

2421

MineContext

MineContext is a proactive, context-aware AI partner that captures screen activity and digital content to automatically generate summaries, to-do lists, and insights.

2422

PixelRAG

PixelRAG is a visual RAG system that renders documents as screenshots instead of parsing them to text, allowing AI to retrieve and reason over visual elements like tables and charts.

2423

helix-db

HelixDB is a graph-vector database built in Rust that consolidates graph, vector, relational, and document data into a single platform for AI memory and knowledge graphs.

2424

honcho

Honcho is memory infrastructure for stateful AI agents that allows them to maintain a persistent, evolving understanding of people and projects through background reasoning and peer-centric representations.

2425

Google Limits Meta's Access to Gemini AI Models Due to Capacity Constraints

Google has restricted Meta's use of Gemini AI models, citing high demand and computing capacity constraints rather than policy-based restrictions.

2426

UltraRAG

A lightweight RAG development framework based on the Model Context Protocol (MCP) that enables low-code orchestration of complex workflows and rapid prototyping via a visual IDE.

2427

trafilatura

A Python package and command-line tool for discovering and extracting clean, structured text and metadata from the web, removing HTML noise to create high-quality datasets.

2428

Current Trends and Zeitgeist in San Francisco

Conversations in San Francisco currently center on the recovery of the city's housing market, the impact of AI and robotics, upcoming international sporting events, and local political tensions.

2429

Suspicious Discontinuities: How Hard Thresholds Incentivize Gaming and Fraud

An exploration of how sharp discontinuities in tax policy, academic grading, and legal thresholds create perverse incentives that lead people to intentionally lose money or manipulate data to cross specific boundaries.

2430

OpenRA playtest-20260222 release notes

OpenRA playtest-2026022 introduces random map generators for Red Alert, Tiberian Dawn, and Dune 2000, alongside a feature-complete Tiberian Dawn HD mod and significant balance overhauls for Dune 2000.

2431

airweave

An open-source context retrieval layer for AI agents and RAG systems that syncs data from 50+ integrations into a unified, LLM-friendly search interface.

2432

vespa

Vespa is a high-performance platform for search, recommendation, and personalization that allows for real-time inferences and data organization using vectors, tensors, and text at scale.

2433

Upsonic

A Python framework for building autonomous and traditional AI agents, featuring secure workspace execution, custom tool integration, and a unified OCR pipeline.

2434

paper-qa

PaperQA2 is an agentic RAG system for scientific literature that provides high-accuracy, grounded answers with in-text citations from PDFs and other document formats.

2435

garden-skills

A curated collection of production-ready skills for AI coding agents (Claude Code, Cursor, Codex) to perform specialized tasks in web design, video production, and image generation.

2436

claude-context

An MCP plugin that adds semantic code search to AI coding agents, allowing them to efficiently retrieve relevant snippets from large codebases using a vector database.

2437

turbovec

A Rust-based vector index with Python bindings that implements Google's TurboQuant algorithm to provide extreme memory compression and fast SIMD-accelerated vector search for RAG applications.

2438

zvec

Zvec is an open-source, in-process vector database that provides low-latency similarity search and hybrid retrieval directly embedded within applications.

2439

coze-studio

Coze Studio is an open-source, low-code visual development platform for creating, debugging, and deploying AI agents, workflows, and AI apps.

2440

DeepTutor

DeepTutor is an agent-native personalized tutoring workspace that integrates tutoring, research, and mastery practice into a single system with shared memory and multi-engine RAG.

2441

kotaemon

An open-source, customizable RAG UI for chatting with documents, featuring hybrid retrieval, multi-modal parsing, and advanced citations.

2442

claude-mem

A persistent memory compression system for Claude Code and other AI CLIs that preserves project context and tool observations across sessions.

2443

opendataloader-pdf

An open-source PDF parser and accessibility tool that extracts structured data (Markdown, JSON) for AI pipelines and automates the creation of Tagged PDFs for accessibility compliance.

2444

ragflow

RAGFlow is an open-source RAG engine that combines deep document understanding with agent capabilities to create grounded, production-ready AI systems from complex unstructured data.

2445

PaddleOCR

A global leading OCR toolkit and document AI engine that converts PDFs and images into structured JSON or Markdown data for LLM-ready applications.

2446

Managing Payment Risks: Is There a 'Bad Employer' List for Unpaid Contracts?

A discussion on Hacker News discusses the lack of a centralized, verified 'bad employer' list for unpaid contracts, highlighting the use of alternative verification methods and legal protections for contractors.

2447

Meta's Legal War on Whistleblower Sarah Wynn-Williams

Meta is using binding arbitration and aggressive non-disparagement clauses to silence former employee Sarah Wynn-Williams, awarding damages exceeding $11 million after she published a memoir detailing institutional misconduct.

2448

DeepSeek DSpark: Accelerating LLM Inference by 60-85%

DeepSeek has open-sourced DSpark, an inference optimization technique that achieves 60-85% faster generation speeds by improving upon speculative decoding.

2449

Normal Computing and the Thermodynamic AI Chip

Thomas Ahle discusses the development of the CN101 thermodynamic chip, which leverages physical noise as a computational resource to solve complex probabilistic workloads and invert matrices.

2450

superglue

An AI-powered tool builder that allows users to create production-grade integrations and tools using natural language, featuring self-healing capabilities for API changes.