🧠 OpenRAG

OpenRAG एउटा RAG प्लेटफर्म हो जसले व्यवसायको सामग्रीलाई सफा र संरचित बनाउँछ, यसलाई हाम्रो आफ्नै सर्भरमा अनुक्रमणिका (इन्डेक्स) गर्छ, र बाह्य एआई प्रणालीहरूद्वारा प्राकृतिक भाषामा प्रश्न गर्न योग्य बनाउँछ एउटा मात्र अनुमति प्राप्त ढोकाबाट।

वेक्टरहरू प्रकाशित गरिन्न; पहुँच खोलिन्छ।

यो एउटा उत्पादनको परिचय हो, वाचाहरूको सूची होइन। तल, आज काम गर्ने भागहरू र अझै रोडम्यापमा रहेका भागहरू छुट्टाछुट्टै रूपमा चिन्ह लगाइएका छन्। रोडम्यापका वस्तुहरूका लागि कुनै मितिहरू वाचा गरिएको छैन; तिनीहरू पूरा भएपछि "आज काम गर्ने" खण्डमा सारिन्छन्।

मुख्य सिद्धान्त

OpenRAG को सम्पूर्ण डिजाइन एउटा वाक्यमा आधारित छ: वेक्टरहरू प्रकाशित गरिन्न, पहुँच खोलिन्छ। बाहिर दिइने कुरा सामग्रीको प्रतिलिपि होइन, तर यसलाई प्रश्न गर्ने अधिकार हो।

⚙️

एउटा वेक्टर एउटा इञ्जिन हो, उत्पादन होइन

एउटा वेक्टर त्यो मोडेलको लागि विशिष्ट हुन्छ जसले यसलाई उत्पादन गरेको छ; अर्को एआईले यसलाई जस्तो छ त्यस्तै प्रयोग गर्न सक्दैन। यसले पाठलाई धेरै गुणासम्म फ्यालाउँछ। त्यसकारण वेक्टरहरू कहिल्यै वितरण गरिन्न।

🔑

जुन खोलिन्छ, त्यो प्रश्न पहुँच हो

एउटा प्रश्न आउँछ; एउटा उत्तर र यो आधारित रहेका स्रोत सामग्रीहरू फर्किन्छन्। पूर्ण सामग्री कहिल्यै प्रतिलिपि गरिन्न, वितरण गरिन्न वा डाउनलोड गरिन्न।

🤝

अनुमति व्यवसायको हो

व्यवसायले कुन सामग्री बाह्य रूपमा प्रश्न गर्न योग्य हुन्छ भन्ने निर्णय गर्छ; सहभागिता वैकल्पिक हुन्छ। अनुमति र सहमतिको तह रोडम्यापमा छ।

यो कसरी काम गर्छ — लक्षित वास्तुकला

यस प्रवाहको सफाइ, संरचना, अनुक्रमणिका (इन्डेक्स) र खोज लिङ्कहरू आज काम गर्छन्। सेल्फ-होस्टेड एम्बेडिङ विथ रि-र्‍यान्किङ, र बाहिरी संसारको लागि एउटा मात्र ढोका, अझै रोडम्यापमा छन्; प्रत्येक चरण तल कुन कुन छ भनेर बताउँछ।

1

स्रोतहरू सङ्कलन गरिन्छ

xloji साइटहरू, अपलोड गरिएका कागजातहरू र क्रल गरिएको वेब पृष्ठहरू सामग्री स्रोतहरू हुन्। (WordPress प्लगइन र सार्वजनिक इञ्जेसन एपीआई रोडम्यापमा छन्।)

2

यो सफा र संरचित गरिन्छ

अव्यवस्थित सामग्रीलाई सफा गरिन्छ, अर्थपूर्ण खण्डहरूमा विभाजन गरिन्छ र RAG प्रयोग गर्न सक्ने रूपमा ल्याइन्छ। यो चरण आज काम गर्छ।

3

यो अनुक्रमणिका (इन्डेक्स) र खोजी गरिन्छ

खण्डहरूलाई वेक्टरहरूमा परिणत गरिन्छ र pgvector मा लेखिन्छ; तिनीहरूलाई HNSW अनुक्रमणिका (इन्डेक्स) र हाइब्रिड खोजी (कीवर्ड + वेक्टर) मार्फत फेला पारिन्छ। यो चरण आज काम गर्छ; रि-र्‍यान्किङ लिङ्क थप्ने रोडम्यापमा छ।

4

यो एउटा मात्र ढोकाबाट पुगिन्छ

बाह्य एआई प्रणालीहरूले प्राकृतिक भाषामा एउटा मात्र अनुमति प्राप्त ढोकाबाट सोध्छन्; जवाफहरू र स्रोत सामग्रीहरू फर्किन्छन्, वेक्टरहरू होइनन्। यो ढोका अझै अस्तित्वमा छैन; यो रोडम्यापमा छ।

आज काम गर्ने भागहरूलाइभ

निम्नहरू आज प्लेटफर्ममा चलिरहेका छन् र पहिले नै xloji.com मा एआई उपकरणहरूद्वारा प्रयोग गरिएका छन्।

🧮

pgvector + HNSW अनुक्रमणिका (इन्डेक्स)

सामग्री खण्डहरू डेटाबेसमा वेक्टरहरूको रूपमा भण्डारण गरिन्छन् र कोसाइन समानतामा HNSW अनुक्रमणिका (इन्डेक्स) संग खोजी गरिन्छ।

🔎

हाइब्रिड खोजी

कीवर्ड खोजी र वेक्टर खोजी सँगै चलिरहेका छन्; दुई परिणाम सूचीहरूलाई एउटा एकल रैंकिंगमा मिलाइन्छ। यसले सटीक शब्द मिलानहरू र अर्थपूर्ण रूपमा नजिकका खण्डहरू पत्ता लगाउँछ।

🧱

मल्टी-टेनेन्ट आइसोलेसन

प्रत्येक व्यवसाय, साइट र लेखकको आफ्नै RAG छ। एउटा प्रश्न केवल यसको आफ्नै टेनेन्टको अनुक्रमणिका (इन्डेक्स) विरुद्ध मात्र चलिरहेको हुन्छ; डेटा कहिल्यै टेनेन्टहरूमा मिसिँदैन।

🧹

कागजात सफाइ र RAG संरचना

अपलोड गरिएका कागजातहरूलाई सफा गरिन्छ, खण्डहरूमा विभाजन गरिन्छ र RAG को लागि संरचित गरिन्छ। प्लेटफर्ममा रहेको डेटा स्ट्रक्चरिङ उपकरण यस पाइपलाइनको प्रयोगकर्ता-फेसिंग अनुहार हो।

🕸️

वेब क्रलिङ

एउटा साइटका पृष्ठहरू क्रल गरिन्छन्, तिनीहरूको सामग्री निकालिन्छ र उस्तै सफाइ र अनुक्रमणिका (इन्डेक्स) पाइपलाइनबाट पार गरिन्छ।

🧩

कस्टम एआई उपकरणहरूमा सार्वजनिक RAG फीड

तपाईंले बनाएको कस्टम एआई उपकरणलाई सार्वजनिक RAG स्रोतबाट फीड गर्न सकिन्छ; उपकरण यसको उत्तरहरूलाई त्यस ज्ञान आधारमा आधारित गर्दछ।

🗄️

फाइल र मिडिया भण्डारण

फाइलहरू र मिडिया वस्तु भण्डारणमा बस्छन् — वेक्टर स्टोरमा होइनन्। RAG रेकर्ड मूल फाइलमा फर्कन्छ।

🏠

हाम्रा आफ्नै सर्भरहरूमा एम्बेडिङ र रिर्‍याङ्किङ

एम्बेडिङ र रिर्‍याङ्किङ अब एउटा छुट्टै सेवामा हाम्रा आफ्नै सर्भरहरूमा चलिरहेको छ। पाठ सर्भर छाडेर नजाँदै भ्याक्टरमा परिणत हुन्छ, र उम्मेदवार परिणामहरू सटीकता सुधार गर्न स्थानीय रिर्‍याङ्करबाट गुज्रन्छ।

✏️

क्वेरी पुन:लेखन

एउटा अस्पष्ट वा अपूर्ण प्रश्न खोजी सुरु हुनु अघि सफा गरिन्छ। प्रयोगकर्ताहरूले "सही" प्रश्न बनाउनु पर्दैन; प्रणालीले क्वेरीलाई खोजीको लागि तयार बनाउँछ।

🎞️

छवि, अडियो र भिडियोबाट पाठ

छविबाट पाठ निकाल्ने र अडियो र भिडियोको प्रतिलेखन सर्भरमा स्थानीय उपकरणहरूसँग सञ्चालन हुन्छ, र परिणामी पाठ उस्तै सफाइ र अनुक्रमणिका पाइपलाइनमा प्रवेश गर्दछ। डेटा सर्भर छाड्दैन।

रोडम्यापमा के छयोजना गरिएको

निम्नहरू अझै अस्तित्वमा छैनन्। तिनीहरू योजना गरिएका र डिजाइन गरिएका वस्तुहरू हुन्; यस पृष्ठले तिनीहरूलाई पहिले नै निर्माण गरिएका रूपमा प्रस्तुत गर्दैन।

🚪

एउटा मात्र ढोका: MCP / API गेटवे

एउटा एकल गेटवे जसले बाह्य एआई प्रणालीहरूले प्राकृतिक भाषामा प्रश्न गर्न सक्छन्। एउटा एआई जसले एक पटक जडान गर्छ, प्रत्येक सहभागी व्यवसाय एउटै ढोकाबाट पुग्छ।

🛡️

अनुमति र सहमति तह

व्यवसायहरू आफ्नै रोजाइमा सामेल हुन्छन्; प्रमाणीकरण, कुञ्जी व्यवस्थापन, दर सीमित गर्ने र दुरुपयोग सुरक्षा सबै यस तहमा छन्।

📡

एआई दृश्यता तह

llms.txt, /.well-known/ र schema.org मार्कअप मार्फत सामग्रीलाई मेसिन-डिस्कभर गर्न योग्य बनाउने, साथै MCP रजिष्ट्रीहरूमा दर्ता गर्ने। यहाँ कुनै जादुई स्वतः-डिस्कभरी छैन; दृश्यता यसरी निर्माण गरिएको छ।

🔌

WordPress प्लगइन

एउटा प्लगइन जसले WordPress साइटलाई यसको सामग्रीलाई OpenRAG मा पठाउन र एउटा एन्डपोइन्ट फिर्ता पाउन दिन्छ जुन एआई प्रणालीहरूले प्रश्न गर्न सक्छन्।

हामी यी वस्तुहरूको लागि मिति दिदैनौं। जब एउटा वस्तु समाप्त हुन्छ, यो "आज काम गर्ने भागहरू" सूचीमा सारिन्छ, र यो पृष्ठ त्यस अनुसार अद्यावधिक गरिन्छ।

🔒 डेटा सर्भर छाड्दैन

OpenRAG को लक्ष्य RAG पाइपलाइनको प्रत्येक चरणलाई हाम्रो आफ्नै सर्भरमा स्थानीय, नि:शुल्क उपकरणहरू प्रयोग गरेर चलाउनु हो: एम्बेडिङ, रि-र्‍यान्किङ, छविहरूबाट पाठ निकाल्ने र अडियो प्रतिलेखन। RAG मा प्रवेश गर्ने कुनै पनि मिडियालाई बाह्य सेवामा पठाइँदैन — यो एउटा गैर-समझौता डिजाइन नियम हो।

ईमानदार स्थिति: यी चरणहरूमध्ये केही आज अझै बाह्य सेवाहरू द्वारा प्रदर्शन गरिन्छ। तिनीहरूलाई घरमै ल्याउनु रोडम्यापमा पहिलो वस्तु हो।

विकास कसरी योजना गरिएको छ

सामग्री बढ्दा वेक्टर खोजी बढी महँगो हुन्छ। OpenRAG को डिजाइनले सुरुमै त्यो लागतलाई तीन तरिकाले सीमित गर्छ।

एउटा ठूलो अनुक्रमणिका (इन्डेक्स) छैन

प्रत्येक व्यवसायको आफ्नै सानो अनुक्रमणिका (इन्डेक्स) हुन्छ र एउटा प्रश्न केवल त्यस अनुक्रमणिका (इन्डेक्स) विरुद्ध मात्र चलिरहेको हुन्छ। कुल सामग्री बढ्दै गए पनि, एउटा प्रश्नले स्क्यान गर्ने क्षेत्र सानै रहन्छ; खोजी लागत त्यस एक व्यवसायको आकारसँग बढ्छ, प्लेटफर्मको कुल आकारसँग होइन।

वेक्टर आकार र परिशुद्धता समायोजन योग्य छ

सानो वेक्टर आयामहरू र थप कम्प्याक्ट संख्या ढाँचाहरू प्रयोग गरेर, उस्तै सामग्रीलाई उल्लेखनीय रूपमा कम स्थान चाहिन्छ। यी विकल्पहरू सुरुमै डिजाइनमा निर्मित छन् ताकि पछि कुनै महँगो माइग्रेसनको आवश्यकता पर्दैन।

कोल्ड टियर

लामो समयसम्म प्रश्न नगरिएको व्यवसायको अनुक्रमणिका (इन्डेक्स) सस्तो स्थायी भण्डारणमा सारिन्छ र यसको लागि प्रश्न आउँदा फिर्ता ल्याइन्छ। त्यसरी, केवल हाल सक्रिय सामग्री मात्र छिटो मेमोरीमा रहन्छ। यो तह रोडम्यापमा छ।

सो Frequently Asked Questions

OpenRAG के हो?

यो एउटा RAG प्लेटफर्म हो जसले व्यवसायको सामग्रीलाई सफा र संरचित बनाउँछ, यसलाई हाम्रो आफ्नै सर्भरमा अनुक्रमणिका (इन्डेक्स) गर्छ, र बाह्य एआई प्रणालीहरूद्वारा प्राकृतिक भाषामा प्रश्न गर्न योग्य बनाउँछ एउटा मात्र अनुमति प्राप्त ढोकाबाट। यो एउटा च्याट उपकरण होइन; यो त्यो उपकरणहरू मुनि चल्ने ज्ञान तह हो।

के तपाईं मेरो वेक्टरहरू वा मेरो सामग्री कसैलाई हस्तान्तरण गर्नुहुन्छ?

होइन। मूल सिद्धान्त यो हो: वेक्टरहरू प्रकाशित गरिन्न, पहुँच खोलिन्छ। बाहिर जाने कुरा सामग्रीको प्रतिलिपि होइन, तर यसलाई प्रश्न गर्ने अधिकार हो; केवल सहायक सामग्रीहरू जवाफको साथ फर्किन्छन्। त्यस माथि, व्यवसायले कुन सामग्री बाह्य रूपमा प्रश्न गर्न योग्य हुन्छ भन्ने निर्णय गर्छ।

आज के वास्तवमा प्रयोग गर्न सकिन्छ?

वेक्टर डेटाबेस र अनुक्रमणिका, हाइब्रिड खोजी, बहु-किराँदार अलगाव, कागजात सफा गर्ने र RAG संरचना, वेब क्रलिङ, अनुकूल AI उपकरणहरूमा सार्वजनिक RAG फिड, र फाइल/मिडिया भण्डारण सबै आज काम गर्छन्। स्व-होस्टेड एम्बेडिङ, पुन: श्रेणीकरण, क्वेरी पुन: लेख्ने, बाह्य ढोका, अनुमति लेयर, दृश्यता लेयर र WordPress प्लगइन अझै सम्म अस्तित्वमा छैनन्।

के मेरो डेटा सर्भर छाड्छ?

RAG पाइपलाइनको प्रत्येक चरण सर्भरमा स्थानीय उपकरणहरूसँग सञ्चालन गर्ने लक्ष्य छ, कुनै पनि मिडिया RAG मा प्रवेश गर्दैन जुन बाह्य सेवामा पठाइन्छ। इमानदार स्थिति: यी चरणहरू मध्ये केही अझै आज बाह्य सेवाहरू द्वारा ह्यान्डल गरिन्छ, र तिनीहरूलाई घर भित्र ल्याउनु कार्यसूचीको पहिलो वस्तु हो।

रोडम्याप सुविधाहरू कहिले तयार हुन्छन्?

हामी मितिहरूको वाचा गर्दैनौं। यही यो पृष्ठको उद्देश्य हो: आज के काम गर्छ र के योजना गरिएको छ भन्ने कुरा छुट्ट्याउन। जब कुनै वस्तु पूरा हुन्छ, यो "आज काम गर्ने भागहरू" सूचीमा सर्छ।

के मेरो सामग्री AI प्रशिक्षणको लागि प्रयोग वा बेचिएको छ?

होइन। आज त्यस्तो कुनै प्रयोग छैन। अनुमति प्राप्त पाठ कोषको विचार रोडम्यापमा सबैभन्दा टाढाको, कानूनी रूपमा भारी वस्तु हो; यो भए पनि, व्यवसायको स्पष्ट सहमति बिना कुनै सामग्री यसको लागि प्रयोग गरिने छैन।

AI पोर्टल मा फर्कनुहोस् अन्य AI उपकरणहरूको लागि।

सो frequently सोधिएका प्रश्नहरू

What is OpenRAG?

OpenRAG is a RAG (Retrieval-Augmented Generation) system and a single-door AI gateway running on xloji.com's own server. When a question arrives, it first finds the relevant content, then generates an answer based on this content. It also serves as a central gateway allowing AI tools outside of xloji.com to access xloji.com's information from a single point.

What does RAG mean and how does OpenRAG implement it?

RAG stands for 'Retrieval-Augmented Generation', meaning it searches for information from relevant sources before an AI generates an answer, then creates a response based on that information. OpenRAG runs these two steps, the search and generation steps, together on its own server. Thus, the answers given by OpenRAG are not random, but based on real content belonging to xloji.com.

What does 'single-door AI gateway' mean?

A single-door AI gateway means that different AI requests and integrations are passed through a single central entry point instead of separate systems. OpenRAG takes on this role: both xloji.com's own 'Ask a Question' feature and AI tools connected from outside receive service through the same OpenRAG infrastructure. This central structure allows AI-related traffic to be managed from a single point.

What difference does it make that OpenRAG runs on its own server?

OpenRAG runs on xloji.com's own server without relying on a third-party cloud service. This means that the searched and processed data remains under xloji.com's own control. As a structure running on its own server, OpenRAG performs search and answer generation operations within its own infrastructure without transferring data to external systems.

Who uses OpenRAG?

OpenRAG is used by three different groups: end-users who visit xloji.com and type into the 'Ask a Question' box, systems that query xloji.com's own RAG infrastructure, and external AI tools that access information about xloji.com via OpenRAG. For the end-user, OpenRAG returns an answer based on xloji.com content to the question asked. For external tools, OpenRAG is a context-free gateway used to retrieve information from xloji.com.

How does the 'Ask a Question' box on xloji.com work with OpenRAG?

A question entered into the 'Ask a Question' box on xloji.com is processed by OpenRAG. OpenRAG first searches for content related to the question within xloji.com's own data, then generates an answer based on the content it finds. This ensures that the answer given to the user is supported by real content from xloji.com.

How do external AI tools obtain information via OpenRAG?

OpenRAG serves as a single-gate passage enabling AI tools outside of xloji.com to access information about xloji.com. These tools can connect via OpenRAG and receive self-contained, understandable answers from xloji.com's content. Therefore, the information provided via OpenRAG is designed to be meaningful on its own, without requiring a separate context.

What are the limitations of OpenRAG?

The answers provided by OpenRAG are limited to the content provided to it; OpenRAG cannot produce a definitive answer about information that is not in the system or has not been added. OpenRAG is designed to remain limited to the content it has, rather than fabricating non-existent information. Therefore, it should be remembered that an answer obtained from OpenRAG is limited to the content that xloji.com currently possesses.

Is OpenRAG a chatbot?

OpenRAG is not a standalone chatbot in the classical sense; it is the RAG and gateway infrastructure that powers xloji.com's 'Ask a Question' feature and external tools' access to information. OpenRAG's core function is to find content corresponding to a question and generate an answer based on that content. As such, OpenRAG operates more as the infrastructure layer behind the answers than as an interface that directly chats with the user.

What is the purpose of OpenRAG for xloji.com?

The purpose of OpenRAG is to enable both end-users and external AI tools to access information belonging to xloji.com through a reliable and self-controlled infrastructure. Running on its own server ensures that this access happens without dependence on third-party systems. Thanks to its single gateway structure, OpenRAG combines different access points in a single central system.

OpenRAG बारे विस्तृत जानकारी

OpenRAG is a RAG (Retrieval-Augmented Generation) system running on xloji.com's own server, and it is also an AI gateway that passes AI-related requests from a single point. The 'Open' in its name indicates that the system offers an open and accessible structure, while 'RAG' refers to its method of operation, which involves first finding relevant data and then generating an answer based on that data. A question typed into the 'Ask a Question' box on xloji.com is first searched for relevant content by OpenRAG, and then an answer is generated based on this content. The same infrastructure also allows AI tools outside of xloji.com to access information belonging to xloji.com through OpenRAG and use this information in their own answers. This structure enables xloji.com to provide services through a single central infrastructure instead of setting up separate AI solutions for its different products.

The basic logic of the RAG approach is that an AI system, instead of relying solely on the general knowledge it was previously trained on, first retrieves up-to-date and accurate content from a source related to the question and then answers in light of this content. OpenRAG operates by combining these two steps – retrieval and generation – on its own server. This approach aims to reduce the tendency of AI models to sometimes generate non-existent information; because the answer is derived from actually found content, not from the model's own memory. In this way, the answers given by OpenRAG are based on xloji.com's own content, not on random or outdated information. Thus, the source of an answer produced by OpenRAG remains traceable as xloji.com's own data. The server belonging to xloji.com means that the searched and processed data is kept under xloji.com's own control without depending on a third-party cloud service.

The 'single-door AI gateway' side of OpenRAG means that different AI requests and integrations are routed through a single central entry point instead of separate systems. In this way, both xloji.com's own 'Ask a Question' feature and external tools connected via OpenRAG use the same infrastructure and the same basic logic. As a central gateway structure, OpenRAG ensures that AI-related traffic passes through a single point; this increases both the consistency of answers and control over the data as a system running on its own server. This single-point approach also allows maintenance and updates to be done in one place instead of in distributed systems. In this respect, OpenRAG is not only a question-and-answer tool, but also an infrastructure layer that unifies all of xloji.com's AI access points.

OpenRAG is used directly or indirectly by three different user groups: end-users who visit xloji.com and type into the 'Ask a Question' box, systems that query xloji.com's own RAG infrastructure, and external AI tools that access information belonging to xloji.com via OpenRAG. All three user groups receive answers fed from the same OpenRAG infrastructure, the same content pool; therefore, a consistent source of information is reached no matter where questions about xloji.com are asked. The answers given by OpenRAG are limited to the content provided to it; that is, a definitive answer should not be expected from OpenRAG about information that is not in the system or has not been added to its server. This limitation also forms the basis of OpenRAG's reliability: instead of fabricating non-existent information, it aims to produce answers limited to and based on the content it has.

OpenRAG — एआई मार्फत तपाईंको सामग्री खोल्नुहोस् एउटा मात्र अनुमति प्राप्त ढोकाबाट | xloji.com