Skip to content
HomeInsights

Insights

Ideas for building
better software.

Engineering perspectives on architecture, product development, AI, and the everyday decisions behind useful software.

From the engineering desk

Browse practical articles or follow the latest technology updates.

Subscribe via RSS →

Curated links from external sources — not 360Softy original articles.

ExternalDatabase
Redis Blog

Prefill vs Decode: LLM Inference Phases Explained

Every LLM request runs in two distinct phases: prefill, where the model reads your prompt in one parallel burst, and decode, where it generates the response one token at a time, each one depending on the last. These two phases have different performan...

Tech DE
Redis BlogRead original
ExternalDatabase
Redis Blog

Long-Term Memory Architectures for AI Agents

Most AI agents start every session from scratch. Without persistent memory, they're stateless responders that reprocess context on every invocation and can't build continuity across interactions. Long-term memory changes that. It gives agents externa...

Tech DE
Redis BlogRead original
ExternalDevOps
Kubernetes Blog

Kubernetes v1.36: Mutable Pod Resources for Suspended Jobs (beta)

Kubernetes v1.36 promotes the ability to modify container resource requests and limits in the pod template of a suspended Job to beta. First introduced as alpha in v1.35, this feature allows queue controllers and cluster administrators to adjust CPU, memory, GPU, and extended resource specifications on a Job while it is suspended, before it starts or resumes running. Why mutable pod resources for suspended Jobs? Batch and machine learning workloads often have resource requirements that are not p

Kubernetes BlogRead original
ExternalCloud
Microsoft Azure Blog

Microsoft Sovereign Private Cloud scales to thousands of nodes with Azure Local

Azure Local scales Microsoft Sovereign Private Cloud, supporting AI and data workloads with full control, compliance, and disconnected operations. The post Microsoft Sovereign Private Cloud scales to thousands of nodes with Azure Local appeared first on Microsoft Azure Blog.

Hybrid + multicloudAIGenerative AI
Microsoft Azure BlogRead original

Let’s start with a conversation

Tell us what you’re working on.

An idea, a challenge, or a system that needs to work better. We’ll help you understand the next step.