llmreference
Baseten API

Using NSQL 350M on Baseten API

Implementation guide · NSQL · NumbersStation

Serverless

Quick Start

  1. 1
    Create an account at Baseten API and generate an API key.
  2. 2
    Use the Baseten API SDK or REST API to call nsql-350m — see the documentation for request format.

Code Examples

See Baseten API documentation for integration details.

About Baseten API

The AI platform offers a comprehensive suite of features designed to streamline the development and deployment of machine learning models. At its core, the platform supports open-source models, allowing developers to leverage existing frameworks and tools for their AI applications. This flexibility is coupled with rapid deployment capabilities, enabling organizations to quickly bring their models into production environments. The platform's architecture is built for scalability, accommodating fluctuating workloads and user demands without compromising performance. A standout feature is its high-speed inference capabilities, crucial for applications that require real-time data processing and decision-making. Cost-effectiveness is a key advantage of the platform, implementing a pay-as-you-go model that minimizes initial investments while optimizing resource utilization. The platform also boasts flexible deployment options, allowing users to deploy models across various environments including cloud, on-premises, or edge devices. This versatility empowers organizations to tailor their deployment strategies to specific needs and existing infrastructure. By combining these features, the platform provides a robust solution that enables businesses to fully harness the potential of AI while maintaining control over costs and deployment logistics.

Baseten is an AI infrastructure platform that provides comprehensive tools for deploying and serving machine learning models efficiently and cost-effectively. The platform offers: 1. Rapid deployment: Users can deploy models in minutes, avoiding complex processes. 2. Support for open-source models: Baseten allows deployment of best-in-class open-source models. 3. Optimized serving: The platform provides optimized serving for custom models. 4. Scalability: Horizontally scalable services enable smooth transition from prototype to production. 5. High-speed inference: Baseten offers fast inference on infrastructure that automatically scales with traffic. 6. Cost-efficiency: The platform includes a scaled-to-zero feature to optimize costs. 7. Flexible deployment options: Models can be run on Baseten's cloud or the user's infrastructure. Baseten aims to simplify the ML deployment process while ensuring performance, scalability, and cost-efficiency for AI builders and developers.

Pricing on Baseten API

Capabilities

No model capability flags are currently sourced.

About NSQL 350M

The NSQL 350M model, developed by Numbers Station, is an open-source large language model tailored for crafting SQL queries from natural language inputs. It belongs to a model family that also includes larger iterations like NSQL 2B and NSQL 6B and is based on Salesforce's CodeGen models. This autoregressive model constructs queries token by token and is recognized for its ability to generate SELECT queries effectively when provided with a table schema and clear instructions. NSQL 350M undergoes initial training using a corpus of general SQL queries and is fine-tuned on a vast dataset consisting of text-to-SQL pairs drawn from diverse public sources. However, while it excels at generating accurate SQL code for moderately complex queries, its efficacy relies heavily on the prompt structure and may falter with highly intricate queries or those outside its training scope. Despite showing strong potential among open-source models, it is yet to match the precision of proprietary models like GPT-4 in tackling complex tasks.

Model Specs

Released2024-02-15
Parameters350M
ArchitectureDecoder Only

Provider

Baseten API
Baseten API

Baseten

San Francisco, California, United States