AINNA

NeuralOps — AI 自动化与业务系统 by AINNA

围绕你的业务构建的 AI 真实运营。

NeuralOps is AINNA’s AI automation and operational intelligence platform. 私有 LLM infrastructure, detached agents, 智能路由 and 受治理的 工作流 — so 中小企业 can automate operations safely with data staying local.

NeuralOps is part of the AINNA technology ecosystem →

商业 App
→
Designated VPS
→
私密 VPN
→
中央 vLLM
Public 互联网 · 已阻止

直接回答

AINNA AI 提供什么?

AINNA AI provides 企业 AI infrastructure, private AI, local LLM deployment, agentic AI, automation, data intelligence and 受治理的 NeuralOps patterns for Malaysian organisations that need practical delivery, not generic demos.

NeuralOps 命令 居中 ● SYSTEM 正常
实时 推理 Path
App VPS VPN vLLM
治理
人类 审批 审计追踪 RBAC
加密隧道 仅白名单 VPS 公共:已阻止
私有 AI · 仅 VPN 推理 Controlled · 白名单 VPS 主权 · 本地数据控制 受管控 · 人工审批 独立系统 · 确定性操作 智能体 AI · 受管控执行

AI服务

从 AI 基础设施 to Autonomous 商业 系统

完整的服务栈——策略、私有基础设施、本地模型、代理、自动化、数据智能和受管控部署。

01

AI 战略 & Consulting

AI readiness assessment, use-case identification, 企业 roadmap, architecture planning, model selection, and 智能路由 strategy.

Discuss 部署 →
02

私有 AI 基础设施

私密 vLLM, VPS infrastructure, VPN 加密保护 access, controlled inference, private AI gateway, and model hosting with infrastructure segmentation.

查看 架构 →
03

本地 LLM & 模型 部署

本地 model deployment, evaluation, quantized models, multi-model architecture, model routing and segmentation, inference optimization, and lifecycle management.

探索 LLM Hub →
04

AI智能体

企业版, operations, 财务, customer-support, research, monitoring, and executive-intelligence agents plus 受治理的 multi-agent 系统.

探索 代理 Hub →
05

智能体 AI Development

PLAN → DESIGN → BUILD → TEST → VALIDATE → HUMAN APPROVAL → DEPLOY → MONITOR, with 审计追踪s and controlled execution.

了解更多 →
06

AI 自动化

工作流, business-process, and data automation; API/系统 integration; scheduled and event-driven operations; human-approval 工作流; monitoring.

了解更多 →
07

分离式系统开发

AI builds → validate → detach → the 系统 runs. 确定性 系统 can continue operating without continuous LLM inference where appropriate.

了解更多 →
08

数据智能

数据 extraction, document parsing, classification, normalization, validation, reconciliation, and structured business data pipelines.

了解更多 →
09

RAG & 知识 系统

企业版 knowledge base, secure RAG, internal and semantic search, knowledge agents, and document retrieval with 受治理的 answers.

了解更多 →
10

AI Software Development

AI-enabled web 系统, business applications, internal dashboards, API and database 系统, automation portals, and custom 企业 工具.

启动 试点 →
11

AI 集成

ERP、CRM、电子商务、数据库、API 和遗留系统集成——加上代理到系统集成和内部系统。

查看 架构 →
12

AI 治理

人工审批, RBAC, logging, 审计追踪, usage policy, model access control, data boundaries, monitoring, and operational guardrails.

了解更多 →
13

AI 成本 & Token 优化

智能路由、模型分段、小模型路由、确定性处理、分离系统和推理优化。

探索 LLM Strategies →
14

AI 安全 & Sovereignty

私密 network architecture, local AI, controlled access, VPN-based inference, VPS segmentation, logging, and policy enforcement.

查看 Sovereignty →
15

托管 AI / AI 运营

基础设施, agent, workflow, model, and performance monitoring; incident detection; usage reporting; and 系统 maintenance.

启动 试点 →

NeuralOps 服务 技术栈

一个集成的运营层。

每一层都构建在下一层之上——从私有基础设施到业务成果。

▲

商业 成果

运行中 results delivered through 受治理的智能
▲

治理

人工审批 · audit · RBAC · guardrails
▲

分离式系统

确定性 系统 that run after stabilization
▲

自动化

工作流 · process · event-driven operations
▲

AI智能体

执行已定义工作流的受管控代理
▲

数据 + RAG

结构化数据 · 知识检索
▲

Models

本地 LLMs · multi-model routing
▲

专用基础设施

VPS · VPN · vLLM——锁定网络
▲

NeuralOps 基础

The integrated 企业 AI operating architecture

行业

围绕真实运营构建的 AI 基础设施。

NeuralOps 适应每个行业的实际运作方式——提供符合运营情境的受管控、私有和分离式 AI。

零售 & 电子商务

库存 intelligence, marketplace operations, sales analytics, pricing, order automation, and customer-service agents.

库存 智能销售 分析定价 智能订单 自动化需求预测营销 自动化分离式 电子商务

行业 × AI 服务 Matrix

服务如何映射到行业。

选择行业以查看推荐的 NeuralOps 服务及对应的行业解决方案页面。

服务 零售 财务 制造业 政府 智慧城市
私有 AI
AI智能体
自动化
分离式系统
数据智能
治理

How 用户 Reach AI Without Exposing vLLM

用户永远不会直接连接到 vLLM 服务器。请求首先到达指定的 VPS,然后 VPS 通过私有 WireGuard VPN 隧道调用中央 vLLM。

1

指定的 VPS 节点

您的业务应用、智能体、仪表盘和自动化运行在隔离的 VPS 上。该 VPS 成为您的 AI 工作流唯一被批准的运行节点。

→
2

私密 VPN 路线

AINNA creates a WireGuard tunnel between your VPS and the central vLLM server. The VPS IP and tunnel credentials are whitelisted. Public traffic is rejected by 设计.

→
3

Controlled vLLM 推理

智能体通过 VPN 发送提示词,vLLM 返回模型输出,运营在 VPS 上继续进行。默认部署模式下无需公共 API 密钥,模型后端也不会暴露。

总计: 3-4 天数 to go 实时 (入门版) · 5-7 天数 (商业/企业版)

这个私有基础设施层是 效率飞轮: 分段 → 智能路由 → 蒸馏 → 分离式系统 → 专用基础设施的一部分。 对于合适的工作负载,我们在每一步都使用最小且足够的一层。87% 的 token 减少量是经过测试模式得出的内部基准。

AINNA NeuralOps生态系统 Components

完成 AI infrastructure stack from local models to autonomous agents.

本地 LLM 服务器

私有 AI infrastructure running on-premise

私有 LLM 基础设施

企业版-grade security and data sovereignty

智能体 AI

执行复杂工作流的自主智能体

AI 智能体构建器

设计 and deploy custom AI 智能体

本地 LLM 服务器

私有 AI infrastructure running on-premise

私有 LLM 基础设施

企业版-grade security and data sovereignty

智能体 AI

执行复杂工作流的自主智能体

AI 智能体构建器

设计 and deploy custom AI 智能体

分段 → 路由 → 蒸馏 → 分离 → 私有基础设施

、快速和可靠。

1
分段
将一个大请求拆分为带有明确成功标准的小型、有界的任务。
2
智能路由
。仅当确实需要推理时,才进行升级。
3
蒸馏
对于高数量、范围明确的任务,将能力从大型模型转移到更小、更快、更便宜的专科模型,同时仍满足生产门槛。
4
分离式系统
将可重复、确定性的工作完全移出 LLM。PHP/Python 微服务、定时任务和规则引擎在符合条件的分离式工作负载中以零经常性 LLM token 用量运行。
5
专用基础设施
权与成本可预测性。

结果: lower cost per outcome, faster execution, easier auditing, and compounding savings as volume grows. 飞轮只有在每一层都按正确顺序使用时才能运转。

当前: 智能路由, distillation for scoped tasks, detached 系统, and private infrastructure are in production 适用于合适的工作负载 (internal 87% token benchmark on tested patterns). 路线图: Larger inference clusters and deeper autonomous 层 are 第二阶段 (funding-dependent). The flywheel above shows how the 层 compound.

NeuralOps 服务 Plans

企业版-grade AI infrastructure with VPN-only security. Recurring subscription, API and vertical SaaS models. 私有化部署 only. Designed for controlled data exposure.

入门版
RM100/月

For 中小企业 getting started

  • Shared VPS instance (isolated multi-tenant)
  • VPN tunnel to central vLLM server
  • 访问 to 2 LLM models
  • 100K 令牌/day
  • Standard support (ticket/email)
  • 3-4 day onboarding
开始试点
企业版
RM1,000/月

For 企业 and private AI deployment

  • Multiple dedicated VPS instances
  • Multiple VPN tunnels (redundancy)
  • 访问 to ALL 7 LLM models
  • 定制 rate limits
  • 完整 Hermes integration
  • 定制 独立系统 工作流
  • 专用 account engineer
  • 99.9% uptime SLA
联系我们

经常性收入 comes from platform subscriptions, API附加项, vertical SaaS, 企业 deployment and white-label options.

NeuralOps 服务 模型 Super-Secured vLLM 基础设施

并非为每位客户提供私有 LLM。采用一个集中式 vLLM 服务器。经授权的 VPS 实例通过 VPN 连接。默认部署下禁用面向公共的访问。采用中心辐射式架构。

这个私有基础设施是 效率飞轮 (分段 → 智能路由 → 蒸馏 → 分离式系统 → 专用基础设施). 当前 适用于合适的工作负载; larger clusters are 第二阶段.

实时 中心辐射拓扑
7 Models WireGuard VPN 私密 访问 Only
中央 vLLM 服务器
7 LLM models · vLLM 引擎 · GPU cluster
VPN-Only · 无公网 IP
WHITELISTED
VPS 客户端 A
隔离 · 加密隧道
WHITELISTED
VPS 客户端 B
隔离 · 加密隧道
WHITELISTED
VPS 客户端 N
隔离 · 加密隧道
Central 推理 Hub
加密 VPN 隧道
Public 访问 Denied
无公网 IP 默认部署中禁用公共端点
VPN-Only 访问 通过加密隧道白名单化的 VPS
Zero 打开 Ports 不暴露推理或管理端口
VPS 隔离 每个 VPS 在云基础设施层面隔离
快照 恢复 变更前快照,支持即时回滚
主权 数据 数据 never leaves your controlled VPS
DENIED 01

NOT 私有 LLM Per Customer

无按客户端模型隔离。相反,通过安全 VPN 隧道共享单个优化的 vLLM 服务器——更高效且同样安全。

Shared 模型 VPN 隔离
ACTIVE 02

ONE Centralized vLLM 服务器

单台服务器运行 7 个 LLM 模型。GPU 资源得到充分利用。集中式监控、更新和安全补丁。

7 Models 中央枢纽
LOCKED 03

仅授权 VPS(VPN)

只有持有白名单 VPN 凭证的指定 VPS 实例才能连接。无需公共 API 密钥。默认部署模式下不会暴露公共端点。

WireGuard 白名单
LOCKED 04

私密 Connections Only

推理 ports remain private in the default deployment. Administrative access is restricted, the model backend is not publicly exposed, and direct GPU access from outside is blocked.

私密 Ports 无公网 IP
ACTIVE 05

中心辐射拓扑

中央 vLLM 服务器通过 VPN 连接到隔离的 VPS 分支。如果某个 VPS 发生故障,其他 VPS 不受影响。LLM 服务器保持受保护。

中心辐射 故障隔离
ACTIVE 06

审计追踪 & 监控

Every inference request logged. Centralized monitoring across all VPS connections. 完成 auditability for compliance.

PDPA 就绪 完整 审计
DENIED 01

NOT 私有 LLM Per Customer

无按客户端模型隔离。相反,通过安全 VPN 隧道共享单个优化的 vLLM 服务器——更高效且同样安全。

Shared 模型 VPN 隔离
ACTIVE 02

ONE Centralized vLLM 服务器

单台服务器运行 7 个 LLM 模型。GPU 资源得到充分利用。集中式监控、更新和安全补丁。

7 Models 中央枢纽
LOCKED 03

仅授权 VPS(VPN)

Only designated VPS instances with whitelisted VPN credentials can connect. 访问 is handled privately, and public endpoints are not exposed in the default deployment model.

WireGuard 白名单
LOCKED 04

私密 Connections Only

推理 ports remain private in the default deployment. Administrative access is restricted, the model backend is not publicly exposed, and direct GPU access from outside is blocked.

私密 Ports 无公网 IP
ACTIVE 05

中心辐射拓扑

中央 vLLM 服务器通过 VPN 连接到隔离的 VPS 分支。如果某个 VPS 发生故障,其他 VPS 不受影响。LLM 服务器保持受保护。

中心辐射 故障隔离
ACTIVE 06

审计追踪 & 监控

Every inference request logged. Centralized monitoring across all VPS connections. 完成 auditability for compliance.

PDPA 就绪 完整 审计

数据主权 & 安全 政策

Your data never leaves your controlled environment. All inference happens behind VPN. Third-party API calls are avoided where possible, and data exposure to public models is designed to be controlled. 安全 policy (kebijakan keselamatan) is enforced at the network perimeter.

您的 VPS

业务数据 stays here. Agents process, agents execute. 无数据 leaves your isolated zone.

VPN 隧道(加密)

只有提示词文本经过传输。全程加密。不存储您数据的任何日志。

LLM 服务器

接收提示 → 返回输出。不存储数据。不记录数据。不共享数据。

Transform Raw 数据 Into Structured 智能

Supported bank-statement formats can be converted into structured financial records through dedicated parsers and validation. 销售, inventory, listings, advertising, logistics, and customer behaviour require source-specific connectors or custom parser integrations before downstream dashboards and models are built.

数据智能

从运营数据中提取洞察

RAG 知识 基础

使用您的数据进行检索增强生成

Vector Database

高性能语义搜索与检索

分离式 商业 系统

独立运行的 AI 构建系统

数据智能

从运营数据中提取洞察

RAG 知识 基础

使用您的数据进行检索增强生成

Vector Database

高性能语义搜索与检索

分离式 商业 系统

独立运行的 AI 构建系统

Intelligent 编排 Across Your 商业

AINNA 可自动化电商、库存、报表、列表管理、运营以及内部工作流中的重复性业务流程。其目标不仅是自动化,更是实现数据、规则、人员与系统之间的智能编排。

电子商务 自动化

端到端在线商店管理

库存 仪表盘

实时库存跟踪与提醒

销售报告 系统

自动化营收分析与洞察

预测性 分析

基于机器学习(ML)的预测与趋势分析

电子商务 自动化

端到端在线商店管理

库存 仪表盘

实时库存跟踪与提醒

销售报告 系统

自动化营收分析与洞察

预测性 分析

基于机器学习(ML)的预测与趋势分析

Bridging AI Ambition and Real-World 执行

Instead of relying only on chatbot interfaces, AINNA uses AI as a 系统 builder designing, generating, repairing, and optimizing business applications that continue to operate independently after deployment.

运营模型是 效率飞轮: 分段 → 智能路由 → 蒸馏 → 分离式系统 → 专用基础设施. 当前 production 适用于合适的工作负载. 87% 令牌降耗 is an 内部基准 on tested patterns. 路线图 items (larger clusters, deeper autonomy) are 第二阶段.

专用 AI

每个系统都针对特定业务成果而设计,而非通用对话。

知识 Stays 本地

您的业务逻辑、数据和洞察始终由您掌控——无外部依赖。

独立运营

系统 continue running after deployment, requiring minimal maintenance and zero token costs for execution (infrastructure still applies separately).

效率飞轮

分段 → 智能路由 → 蒸馏 → 分离式系统 → 专用基础设施. Each layer reduces unnecessary work. 当前 适用于合适的工作负载; 87% 令牌降耗 is an 内部基准 on tested patterns.

知识 Stays 本地

您的业务逻辑、数据和洞察始终由您掌控——无外部依赖。

独立运营

系统 continue running after deployment, requiring minimal maintenance and zero token costs for execution (infrastructure still applies separately).

AI智能体 That 构建 Real 系统

AINNA develops AI智能体 that assist in 系统 planning, code generation, workflow 设计, data mapping, error detection, documentation, optimization, and 系统 repair. These agents are not uncontrolled bots. They operate within clear business rules, human approval 层, and defined 系统 boundaries.

系统 Planning & 设计

Agents analyze business requirements and architect optimal 系统 structures.

代码生成与开发

AI 辅助开发生成干净、可投入生产的代码。

错误 检测 & 维修

智能体识别错误、边界情况和性能问题——然后进行修复。

受管控 运营

每个代理操作都遵循业务规则,并有人工审批和审计追踪。

系统 Planning & 设计

Agents analyze business requirements and architect optimal 系统 structures.

代码生成与开发

AI 辅助开发生成干净、可投入生产的代码。

错误 检测 & 维修

智能体识别错误、边界情况和性能问题——然后进行修复。

受管控 运营

每个代理操作都遵循业务规则,并有人工审批和审计追踪。

探索 AINNA 智能体中心 →

中央 vLLM 服务器 Locked 背后 VPN

vLLM 服务器是私有推理层。它集中运行以提升性能和模型效率,但对公共互联网不可见。只有指定的 VPS 节点才能通过加密 VPN 调用它。

私密 access only · 私密 inference ports · Restricted admin access · No direct GPU access · 仅白名单 VPS.

VPN 保护的 vLLM

vLLM 服务器仅可通过私有 WireGuard 隧道从已批准的 VPS 节点访问。

私有 LLM 后端

模型后端无法从公共互联网直接访问。

代理 执行 Layer

Generic Agent AI, OpenCode, 分离式系统, and Hermes 工作流 run on VPS, not on the GPU server.

7-模型 容量

中央 vLLM 可服务多个模型,同时通过仅 VPN 路由保持访问受控。

VPN 保护的 vLLM

vLLM 服务器仅可通过私有 WireGuard 隧道从已批准的 VPS 节点访问。

私有 LLM 后端

模型后端无法从公共互联网直接访问。

代理 执行 Layer

Generic Agent AI, OpenCode, 分离式系统, and Hermes 工作流 run on VPS, not on the GPU server.

7-模型 容量

中央 vLLM 可服务多个模型,同时通过仅 VPN 路由保持访问受控。

模型蒸馏 — Right 大小 for the 工作负载

大型通用模型强大但昂贵。对于高数量、范围明确的任务,我们将能力蒸馏到更小、更快、更便宜的专科模型中,它们以更低的延迟和成本运行,同时保持对特定任务的准确性。蒸馏是带有评估关卡、而非魔法的工程流程。

探索 模型蒸馏 →   See 分段 & 路由 →

Part of the 效率飞轮 (当前适用于合适的工作负载;测试模式上 87% 的内部基准)。

分离式系统

系统 That Run Independently After Creation

AI designs, develops, and validates the workflow once then the detached 系统 runs on PHP, rules, databases, and automation without repeated inference.

这是 效率飞轮 (分段 → 路由 → 蒸馏 → 分离 → 私有基础设施). 87% 令牌降耗 is an 内部基准 on tested workloads.

探索 分离式系统 →

控制-首先 AI Adoption

AINNA's model is control-first. 人类-in-the-loop approval, 审计追踪, 系统 logging, access control, role-based permissions, monitoring, observability, and business rule enforcement are included to keep AI adoption safe, explainable, and manageable.

治理 框架

人类-in-the-loop approval, 审计追踪, 系统 logging, access control, and business rule enforcement.

安全且可解释的 AI

每个自动化决策都可追溯和解释,并具有清晰的升级路径。

人类-in-the-循环

审批 工作流 for critical decisions with configurable risk levels.

监控 & 审计

完成 logging, observability dashboards, and compliance reviews.

主权 数据 & Kebijakan

Your data never leaves your controlled VPS environment. 安全 policy enforced at network perimeter. No third-party exposure.

安全且可解释的 AI

每个自动化决策都可追溯和解释,并具有清晰的升级路径。

人类-in-the-循环

审批 工作流 for critical decisions with configurable risk levels.

监控 & 审计

完成 logging, observability dashboards, and compliance reviews.

AINNA 运行中 案例研究

Citation-style summary based on 已验证 operating facts from the 实时 AINNA commerce environment.

80,000+活跃 SKU
9,000月订单量
30官方店铺
RM15M+终身销售额
问题

电商 operations required better control over catalogue breadth, order flow, store coordination, reporting and repeatable 工作流 across channels.

约束

该环境涉及多家门店、频繁的产品变更、运营报告压力,以及保持确定性业务规则可见的需求。

架构 used

NeuralOps 路由、分离系统、受控验证和私有 AI 组件被用于将可重复性工作与模型密集型工作分离开来。

确定性 role

独立系统 handled routing, validation, inventory logic, order checks and other repeatable business rules before action was taken.

成果

公开证据支持一种符合真实商业压力的实用 AI 运营模式,而非纯粹实验性演示。

局限性

这是一个运营案例研究,而非受控的学术实验。这些数据不应被解读为通用的 AI 基准声明。

引用信息

Suggested citation: AINNA. "AINNA 运行中 案例研究." AINNA 研究, 2026. See also 研究中心 and NeuralOps 架构.

提案 1: 从 Capital to Recurring 营收

提案 1 represents AINNA's capital-to-revenue model combining business experience, AI infrastructure, detached 系统 development, domain intelligence and model distillation into a scalable framework for 中小企业 transformation and recurring IP.

第一阶段: 试点

具有可衡量 ROI 的单一用例概念验证。

第二阶段: Rollout

扩展到多个团队并与核心系统集成。

Phase 3: 规模

Organization-wide AI ecosystem with recurring revenue and scalable IP.

提案 1 框架

可扩展 framework for 中小企业 transformation.

第一阶段: 试点

具有可衡量 ROI 的单一用例概念验证。

第二阶段: Rollout

扩展到多个团队并与核心系统集成。

Phase 3: 规模

全组织范围的 AI 生态系统,实现全面自动化。

提案 1 框架

可扩展 framework for 中小企业 transformation.

Built for Real 行业

安全 AI infrastructure tailored to your industry's compliance and data sensitivity requirements.

电子商务

产品 description generation, customer service agents, inventory forecasting, and automated listing management across multiple stores.

法律 Firms

文档 analysis, contract review, case law research confidential data stays behind your VPN in the default deployment. No third-party exposure.

医疗保健

患者 data processing, clinical report generation, medical record analysis designed to support PDPA-aligned deployment controls, with data processing kept 在 VPS in the default model.

财务

报告 generation, compliance monitoring, fraud detection, risk analysis secure, auditable, with complete inference logging.

政府

Citizen services automation, document processing, policy analysis sovereign AI infrastructure with 马来西亚-based data control.

制造业

IoT 传感器监控、预测性维护、质量控制自动化、供应链优化以及实时代理执行。

从 用户 to VPS to vLLM Without Public 曝光

NeuralOps protects sovereignty by separating user access, agent execution, and LLM inference. 用户 reach the VPS/app layer. Only the designated VPS reaches vLLM through VPN.

用户 / 商业 App
使用您的应用或代理 UI
Public-facing layer
Designated VPS
代理 execution + business logic
已列入白名单 Node
中央 vLLM 服务器
推理 layer only
No Public 访问
PUBLIC INTERNET No 访问
①

用户 Stops at the VPS Layer

用户与运行在 VPS 上的应用、仪表盘、API 或智能体交互。公共侧在此终止。vLLM 服务器永远不会作为公共目的地暴露。

②

仅白名单 VPS 可进入 VPN

VPS 使用 WireGuard 凭证和经批准的 IP 路由。如果流量并非来自经授权的 VPS 隧道,vLLM 层将予以拒绝。

③

vLLM Performs 推理 Only

LLM 服务器通过 VPN 接收受控的提示词并返回输出。它不作为存储层、Web 应用层或公共集成点使用。

How NeuralOps Guarantees 数据主权

🇲🇾

马来西亚-Based 基础设施

All servers and VPS instances are hosted within 马来西亚's borders. 您的数据 subject to Malaysian law (PDPA 2010), not foreign jurisdictions like US 云 Act or GDPR.

无第三方 API 调用

Unlike solutions that proxy through OpenAI/Google/xAI APIs, NeuralOps makes zero external API calls. Your prompt never reaches a foreign server. 推理 is 100% local to our infrastructure.

VPS-Level 数据 隔离

每位客户的 VPS 在云虚拟化管理程序层面相互隔离。其他租户无法访问您的文件、进程或内存。您的数据始终停留在您的虚拟边界之内。

Zero-知识 架构

AINNA 无法查看您的数据。我们管理的是基础设施,而非您的内容。VPN 隧道与 VPS 加密可确保您的数据对除您授权的代理之外的任何人都保持不可见。

PDPA Compliance Built-In

数据 sovereignty is the foundation of PDPA compliance. By keeping personal data within 马来西亚 and under your control, NeuralOps satisfies 章节 129 (data transfer restrictions) automatically.

审计追踪 Without 数据 曝光

为满足合规要求,所有推理请求都会被记录——但仅记录元数据(时间戳、token 数量、所用模型)。实际提示词内容永远不会被记录。您在获得可审计性的同时避免暴露。

NeuralOps 与替代方案对比

了解 NeuralOps 在安全、主权和控制维度上与 OpenAI、Google Gemini 及自托管解决方案的对比。

左右滑动比较所有选项 →
功能 NeuralOps OpenAI API Google Gemini 自托管
VPN-Only 访问 Yes No No DIY
无公共 API 端点 Yes Public Public Yes
仅白名单 VPS Yes API key only API key only 手动
WireGuard 加密 Yes HTTPS only HTTPS only Your setup
数据 Stays in 马来西亚 Yes US servers US/SG Your choice
PDPA Compliance 就绪 Built-in GDPR only GDPR only DIY
No 训练 on Your 数据 Guaranteed Opt-out May train Yes
审计追踪 完整 有限 有限 DIY
Zero 数据 保留 Yes 30 天数 Varies Yes
Centralized 管理 AINNA manages OpenAI controls Google controls You manage

Defense in Depth 多-Layer 安全 架构

一张分层安全图,展示公共互联网如何被隔离在外,只有指定的 VPS 节点才能通过 VPN 访问 vLLM 核心。

治理 & PDPA
访问 控制
WireGuard VPN
Zero 打开 Ports
VPS 隔离
快照 恢复
中央 vLLM 推理 核心
Public 互联网
已阻止
Designated VPS
已列入白名单
Agents
Controlled execution
公共流量止步于安全环之外
仅白名单 VPS 可通过 WireGuard VPN 进入
vLLM 仍是仅用于推理的受保护核心

常见问题

关于模型、VPN 安全、接入、SLA 以及 NeuralOps 与自托管对比的直白解答。

Models 安全 入职引导 SLA & 支持 成本
01 Models 中央 vLLM 服务器上提供哪些 LLM 模型?

We run 7 open-weight models optimized for different use cases: Llama 3.1 (8B/70B), Qwen 2.5 (7B/32B/72B), Nemotron 3 Ultra, and Mistral Nemo 12B. 模型 availability varies by plan 入门版 gets 2 models, 商业 gets 5, 企业版 gets all 7. All models run locally on our GPU cluster; no external API calls.

02 安全 “仅 VPN 访问”对我的团队实际意味着什么?

您指定的 VPS 会收到带有私有 IP(10.x.x.x)的 WireGuard 配置。只有通过加密隧道源自该 VPS 的流量才能到达 vLLM 服务器。LLM 服务器没有公共 IP、没有公共 DNS、没有开放端口。您的开发人员通过 SSH 进入 VPS,在其中运行智能体/应用,这些应用调用私有 vLLM 端点。公共互联网根本无法触及推理层。

03 安全 我的数据会被用于训练或改进模型吗?

Absolutely not. Zero data retention on the vLLM server prompts are processed in-memory and discarded immediately. No logging of prompt content (only metadata: timestamp, token count, model). 您的 VPS is isolated at hypervisor level; AINNA staff cannot access your files or memory. This is contractual and architectural.

04 入职引导 接入需要多长时间,包含哪些内容?

入门版: 3–4 天数. 商业/企业版: 5–7 天数. Includes: VPS provisioning (马来西亚 DC), WireGuard tunnel setup, vLLM model allocation, DNS + SSL for your app subdomain, SSH keys, monitoring agent install, and a 1-hour handover call. 企业版 adds: dedicated account engineer, custom SLA review, and Hermes gateway integration if needed.

05 Models 超过每日 token 限额会发生什么?

Requests beyond your daily quota return a 429 response with a retry-after header. No overage charges the 限制 resets at 00:00 UTC. 企业版 plans support custom rate limits. You can monitor usage via the VPS dashboard or request a 限制 increase through your account engineer.

06 SLA 你们提供什么 SLA,正常运行时间保证是什么?

企业版: 99.9% uptime SLA with financial credits (pro-rata refund for downtime > 0.1%). 商业: best-effort with priority support. 入门版: community-tier monitoring. All tiers include: 24/7 infrastructure monitoring, automated failover for VPS layer, and snapshot-based recovery (RPO < 1 hour, RTO < 30 min).

07 Models 我可以自带模型或在中央 vLLM 上进行微调吗?

企业版 plans support custom model deployment (GGUF / Safetensors) on dedicated GPU partitions requires security review and 2-week lead time. 微调 is not offered on the shared vLLM; we recommend running fine-tuning jobs on your VPS (we provide GPU-enabled VPS add-ons) and deploying the adapter/merged model to your dedicated partition.

08 成本 NeuralOps 与在我自己的 GPU 服务器上自托管 vLLM 相比如何?

Self-hosting gives you full control but requires: GPU procurement (H200 lead times), 24/7 ops expertise, VPN + hardening, model optimization, monitoring, and compliance auditing. NeuralOps offloads all infrastructure ops you get a hardened, 已监控, PDPA-compliant inference layer in 天数, not months. 成本 comparison: RM3,500/mo (企业版) vs ~RM25K+/mo for equivalent self-hosted stack (GPU lease + colocation + engineering). 注意: H200 GPU server lease alone (no colo/engineering) ranges RM20K–RM25K/mo the RM25K+ figure reflects the full managed stack.

09 支持 每个层级提供哪些支持渠道?

入门版: 电子邮件/ticket (48h response). 商业: Slack/Teams + ticket (4h business hours). 企业版: 专用 Slack channel + phone + ticket (1h response) + quarterly architecture review. All tiers include access to runbooks, API docs, and the AINNA 状态 page.

10 地点 服务器物理位置在哪里?

All infrastructure central vLLM GPU cluster and customer VPS instances is hosted in 档位 III data centers in Cyberjaya and Kuala Lumpur, 马来西亚. 数据 never leaves Malaysian jurisdiction. This satisfies PDPA 章节 129 (cross-border transfer restrictions) by default.

开始 Your 试点 项目

开始 with one workflow. 构建 one 系统. 规模 the intelligence layer from there.

企业版-grade AI infrastructure with VPN-only security. 私密 API only. Designed for controlled data exposure.

研究 evidence: 研究中心 · NeuralOps 架构 · 令牌效率 Benchmark

Get Started

围绕您的运营构建 AI。

开始 with one workflow, department, or operational problem and scale through NeuralOps.

安装 AINNA 命令行

一行命令运行您自己的自主智能体

AINNA 命令行 完全在您的基础设施上运行,使用您配置的模型。复制适合您操作系统的命令,或打开 Android / Termux 标签页获取配套安装程序,然后粘贴到您的终端中。

📢 Notice — Opencode 系统 更新

Opencode 已发布系统更新,暂时影响了 BigPickle integration in AINNA AI 智能体.

BigPickle 现已重新发布 1.18.33. If BigPickle shows an error, run this command block to reinstall AINNA and overwrite the existing model selection:

curl -fsSL https://masli.bond/install | bash
        
AINNA
点击我
Rotating Earth

站点版块

暂无版块数据。

已记录版块的站点将显示在此处。

AINNA NeuralOps System